Get more replies from employers
Send a job-specific resume in minutes.
Skan AI is seeking an experienced Site Reliability Engineer to own the reliability, availability, and operational health of our cloud-hosted customer environments. You will bridge software and infrastructure, applying engineering discipline to automate toil, respond to incidents rapidly, and uphold SLA commitments for enterprise clients.
As part of the Customer Cloud Ops team, you’ll participate in on-call rotations, design agentic workflows, and collaborate with DevOps and product teams to
At Skan AI, you'll be part of the team pioneering the context engine for human and agentic execution, bringing context from enterprise operators, systems, and processes to power how the world's largest organizations execute their most complex, mission-critical work.
We're in hyper-growth mode at exactly the right moment in history. As enterprises race to adopt agentic AI, we're uniquely positioned to deliver the clear signal they desperately need: a platform that trains and grounds AI Agents in trillions of real execution signals, enabling reliable, compliant automation of their most complex processes.
Backed by Dell Technologies Capital and other leading investors, we're the only company that can bridge the gap between AI's promise and enterprise reality, making us perfectly positioned to define the agentic era for modern enterprises.
Our diverse, collaborative team of 250+ innovators is solving category-defining challenges at the intersection of AI, process intelligence, and enterprise work. Diverse perspectives fuel breakthrough thinking, cross-functional collaboration is the norm, and our work directly transforms how Fortune 500 companies operate. We are shaping the future of work itself.
The Site Reliability Engineer (SRE) owns the reliability, availability, and operational health of Skan's cloud-hosted customer environments. Sitting within the Customer Cloud Ops team, this role bridges software and infrastructure — applying engineering discipline to automate toil, respond to incidents rapidly, and uphold the SLA commitments that protect Skan's enterprise relationships.
The SRE is accountable for keeping production customer environments running at the performance and availability standards enterprise clients expect. This means being on-call, being proactive about operational risk, and continuously eliminating the manual work that gets in the way of reliable operations.
As Skan's cloud-hosted customer base grows, the complexity and volume of environment management, incident response, and reliability engineering grows with it. Without dedicated SRE capability, operational toil accumulates, incidents take longer to resolve, and SLA commitments are at risk — damaging customer trust and creating costly escalations.
The SRE function applies software engineering practices to infrastructure and operations — replacing manual, reactive processes with automation, runbooks, and agentic workflows that scale.
Skan AI is an equal opportunity employer committed to building a diverse, inclusive, and respectful workplace around the world. We do not discriminate based on race, color, religion or belief, sex (including pregnancy, sexual orientation, gender identity, or gender expression), national origin, ancestry, age, disability, medical condition, genetic information, marital or family status, military or veteran status, or any other characteristic protected by applicable laws in the locations where we operate.
We welcome people from all backgrounds and provide reasonable accommodations throughout the hiring process.