Turn this role into an interview — a resume and cover letter built around what this employer wants.
Stryker Group in Bengaluru, India is seeking an experienced Site Reliability Engineer to own production systems, lead incident response, and drive reliability improvements across cloud infrastructure.
You will design and manage scalable AWS-based infrastructure using Terraform, Kubernetes (EKS), and implement GitOps with GitLab CI and ArgoCD while collaborating with global teams and ensuring SOC2/ISO27001 alignment.
Own and maintain highly available production systems, lead incident response (P1/P2), conduct RCA/PIRs, and drive improvements to reliability, performance, and operational excellence.
Design, build, and manage scalable cloud infrastructure on AWS using Terraform, with strong ownership of Kubernetes (EKS), networking, security, and platform resilience.
Develop and optimize CI/CD and GitOps pipelines using GitLab CI and ArgoCD, while automating operational processes to improve efficiency and consistency.
Manage observability and on-call operations through tools such as PagerDuty/Zenduty, Prometheus, Grafana, ELK, and Datadog, ensuring actionable monitoring and effective alert management.
Collaborate with global engineering, security, and product teams, contribute to cloud architecture and compliance initiatives (SOC2, ISO27001), create operational documentation, and mentor team members.
4–8 years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles.
Strong hands-on experience with AWS, Terraform, Kubernetes, andArgoCDin production environments.
Experience operating EKS-based platforms, including networking, scaling, monitoring, and troubleshooting.
Strong knowledge of CI/CD,GitOps, and automation practices, with hands-on use of GitLab CI andArgoCD.
Experience managing production systems in a 24×7 environment, including incident response and on-call practices.
Solid Linux and cloud networking background.
Experience with observability tools such as ELK and Prometheus.
Strong scripting skills in Python, Bash, or Go.
Engineering degree in computer science or equivalent.
Cloud certifications such asAWSSysOps/ DevOps Engineer, or CKA/CKAD.
Exposure to ITSM or change management processes in regulated industries (healthcare, fintech, or similar).