A complete application in a minute — tailored resume and cover letter, ready to send.
Get past ATS filters
Job summary
A technology firm is seeking a Lead Site Reliability Engineer to design and implement automated infrastructure and manage Kubernetes workloads. The role involves refining CI/CD pipelines and leading incident response efforts, requiring expertise in Terraform, Prometheus, and Grafana. With over 12 years of industry experience, the ideal candidate will ensure architectural scalability and reliability, emphasizing security and system resilience in a cloud-native environment. This is a pivotal role in delivering highly available and performant systems.
Qualifications
Minimum of 12 years of industry experience in software and systems engineering.
Proven ability in designing automated infrastructure.
Experience with monitoring and observability solutions.
Responsibilities
Design and implement automated infrastructure.
Manage containerized workloads within Kubernetes.
Refine CI/CD pipelines for seamless code deployment.
Lead incident response and post-mortem analyses.
Skills
Terraform
Kubernetes
CI/CD pipelines
Prometheus
Grafana
Automation
Security mindset
System resilience
Job description
A technology firm is seeking a Lead Site Reliability Engineer to design and implement automated infrastructure and manage Kubernetes workloads. The role involves refining CI/CD pipelines and leading incident response efforts, requiring expertise in Terraform, Prometheus, and Grafana. With over 12 years of industry experience, the ideal candidate will ensure architectural scalability and reliability, emphasizing security and system resilience in a cloud-native environment. This is a pivotal role in delivering highly available and performant systems.