Turn this role into an interview — a resume and cover letter built around what this employer wants.
Get past ATS filters
Job summary
A pioneering open source technology company is seeking a Senior Cloud Infrastructure Engineer to design, maintain, and optimize large-scale systems for AI workloads. The role involves managing GPU environments, collaborating on infrastructure design, and driving best practices within a fast-growing team. Candidates should have substantial experience with Python, Kubernetes, Terraform, and Ansible.
Qualifications
5+ years experience as an Infrastructure Engineer or Site Reliability Engineer building and operating large-scale distributed systems.
Skilled in Python and comfortable with Terraform and Ansible.
Familiar with Kubernetes and related tools like FluxCD, Prometheus, and Grafana.
Responsibilities
Design, build, and maintain core infrastructure for AI workloads.
Manage and automate GPU compute clusters using Python, Kubernetes, Terraform, and Ansible.
Collaborate with core engineers on infrastructure for new features.
Skills
Python
Kubernetes
Terraform
Ansible
FluxCD
Prometheus
Grafana
Job description
A pioneering open source technology company is seeking a Senior Cloud Infrastructure Engineer to design, maintain, and optimize large-scale systems for AI workloads. The role involves managing GPU environments, collaborating on infrastructure design, and driving best practices within a fast-growing team. Candidates should have substantial experience with Python, Kubernetes, Terraform, and Ansible.