Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Get past ATS filters
Benefits offered by this job
Medical, dental, vision benefits
Unlimited PTO
Generous paid parental leave
401(k) plan with company match
Job summary
A leading Voice AI platform in the United States is searching for an experienced Site Reliability Engineer to build and operate its hybrid infrastructure. You will be responsible for architecting and maintaining the computing platform using Kubernetes and AWS, developing infrastructure as code with Terraform, and optimizing AI/ML job scheduling systems. The ideal candidate has 5+ years of experience in SRE or DevOps and a strong background in managing production infrastructure. This role offers competitive benefits and the opportunity to work on cutting-edge AI technologies.
Qualifications
5+ years of experience in Platform Engineering, DevOps, or Site Reliability Engineering.
Hands-on experience building and managing production infrastructure with Terraform.
Expert-level knowledge of Kubernetes in large-scale environments.
Responsibilities
Architect and maintain core computing platform using Kubernetes.
Develop and manage infrastructure as code with Terraform.
Design and optimize AI/ML job scheduling systems.
Skills
Platform Engineering
DevOps
Site Reliability Engineering
Kubernetes
Terraform
Python
Go
Bash
Tools
Slurm
AWS
CI/CD tools
Job description
A leading Voice AI platform in the United States is searching for an experienced Site Reliability Engineer to build and operate its hybrid infrastructure. You will be responsible for architecting and maintaining the computing platform using Kubernetes and AWS, developing infrastructure as code with Terraform, and optimizing AI/ML job scheduling systems. The ideal candidate has 5+ years of experience in SRE or DevOps and a strong background in managing production infrastructure. This role offers competitive benefits and the opportunity to work on cutting-edge AI technologies.