Get more replies from employers
Send a job-specific resume in minutes.
Litmus Automation Inc. is hiring a Site Reliability Engineer to own the day-to-day reliability, security, and performance of our Azure-hosted UNS and LEM environments deployed for a strategic enterprise customer.
This role covers provisioning and hardening, deployment support, monitoring, incident response, and long-term operation against a 99.9% uptime commitment. You will work with AKS, Azure networking, managed databases, identity federation, and MQTT-based data pipelines across multiple
Litmus is building the data foundation that powers industrial AI.
AI doesn’t work without real-world, contextualized data - Litmus makes that data usable. As AI adoption accelerates, most industrial environments still can’t access or use their operational data. We solve that gap.
We’re a growth-stage software company helping manufacturers access, structure, and use real-time data from machines, systems, and sensors at the edge. Our platform sits at the intersection of edge computing, AI, and industrial operations, enabling some of the world’s largest companies to run operations in real time, reduce downtime, and optimize production.
Backed by leading investors and trusted by global manufacturers and partners like Google, Microsoft, Dell, Oracle, and Mitsubishi, Litmus is powering the shift toward software-defined manufacturing.
AI is moving beyond the cloud and into the physical world. At Litmus, you’ll build the infrastructure that enables real-time data to power AI and machine learning systems in production environments.
Most AI systems fail without access to real-world data. You’ll build the layer that makes them viable in production. We solve challenges at the intersection of distributed systems, real-time data, and industrial constraints — where reliability, scale, and performance are non-negotiable.
You’ll work on systems used by real customers in production, with direct impact on product and company trajectory. As a scaling company, we move quickly. You’ll have ownership, visibility, and the ability to shape both product and company as we scale.
We’re building a team that holds a high bar and pushes each other to improve. You’ll work alongside experienced operators, engineers, and leaders who have done this before and are building again at scale. We hire people who take ownership, move quickly, and care about outcomes. No passengers.
At Litmus, the team is collaborative, curious, and low ego. People are scrappy, take ownership, and look for ways to make an impact. We value empathy just as much as execution, whether that’s in how we build, how we communicate, or how we support each other.
We’re a growing company, so things move quickly and not everything is perfectly defined. If you enjoy figuring things out, working closely with others, and making steady progress, you’ll do well here.
Litmus Automation is hiring a Site Reliability Engineer to own the day-to-day reliability, security, and performance of our Azure-hosted Litmus Unified Namespace (UNS) and Litmus Edge Manager (LEM) environment, deployed for a strategic enterprise customer. This role covers the full lifecycle: initial environment provisioning and hardening, deployment support, ongoing monitoring and incident response, and long-term operation against a 99.9% uptime commitment. You will work directly with customer-facing infrastructure spanning AKS, Azure networking, managed databases, identity federation, and MQTT-based data pipelines connecting multiple sites. As this is an SRE role held to a 99.9% uptime commitment, it includes participation in an on-call rotation to ensure continuous coverage and rapid incident response outside of standard working hours.