A dynamic tech firm located in San Francisco is seeking a Site Reliability Engineer to enhance operational health across their production systems. This high-impact role demands expertise in AWS and strong programming skills. You will manage production systems' reliability and lead incident response efforts to prevent issues, all while contributing to the scalability and efficiency of their services. Ideal candidates will have 5+ years of relevant experience and a passion for leveraging technology to drive outcomes.
Qualifications
5+ years in Site Reliability Engineering, DevOps, or systems engineering roles.
Strong programming skills in Python, Go, or TypeScript/Node.js.
Experience with infrastructure-as-code and observability solutions.
Responsibilities
Own reliability, availability, and performance of Gamma's production systems.
Build observability infrastructure with metrics and logging.
Lead incident response and drive systemic improvements.
Skills
Site Reliability Engineering
AWS expertise
Python
Go
TypeScript/Node.js
Terraform
CloudFormation
Docker
Kubernetes
Education
Bachelor's degree in relevant field
Tools
CloudFormation
Docker
Kubernetes
Kafka
Job description
A dynamic tech firm located in San Francisco is seeking a Site Reliability Engineer to enhance operational health across their production systems. This high-impact role demands expertise in AWS and strong programming skills. You will manage production systems' reliability and lead incident response efforts to prevent issues, all while contributing to the scalability and efficiency of their services. Ideal candidates will have 5+ years of relevant experience and a passion for leveraging technology to drive outcomes.