An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Get past ATS filters
Job summary
A pioneering open source technology company in San Francisco is seeking a Senior Cloud Infrastructure Engineer to design and manage large-scale distributed systems. The ideal candidate has over 5 years of experience in infrastructure engineering, is skilled in tools like Python, Kubernetes, and Terraform, and thrives in a fast-paced startup environment that values autonomy. This is an on-site position.
Qualifications
5+ years experience as an Infrastructure Engineer or Site Reliability Engineer building and operating large-scale distributed systems.
Skilled in Python and comfortable working with infrastructure-as-code tools such as Terraform and Ansible.
Familiar with container orchestration systems such as Kubernetes and related tooling.
Responsibilities
Design, build, and maintain the core infrastructure that powers AI workloads at scale.
Manage and automate GPU compute clusters using tools such as Python, Kubernetes, Terraform, and Ansible.
Collaborate closely with core engineers to design infrastructure for new features and systems.
Skills
Python
Kubernetes
Terraform
Ansible
Prometheus
Grafana
FluxCD
Infrastructure
Job description
A pioneering open source technology company in San Francisco is seeking a Senior Cloud Infrastructure Engineer to design and manage large-scale distributed systems. The ideal candidate has over 5 years of experience in infrastructure engineering, is skilled in tools like Python, Kubernetes, and Terraform, and thrives in a fast-paced startup environment that values autonomy. This is an on-site position.