Turn this role into an interview — a resume and cover letter built around what this employer wants.
BuildxPartners in Bengaluru is seeking a Senior Site Reliability Engineer to manage reliability, scalability, and performance of production systems in a hybrid/multi-cloud environment.
You will own SLI/SLOs, manage Kubernetes-based platforms (EKS/OpenShift) across AWS and IBM Cloud, and drive incident response and postmortems while mentoring juniors.
Bangalore South, India | Posted on 15/09/2026
BuildxPartners is a global talent solutions firm delivering end-to-end recruitment and workforce solutions across industries and geographies.
→ BuildxAlpha – Executive & Leadership Search Focused on C-suite, board, and global executive hiring.
→ BuildxSigma – Comprehensive Talent Across Levels Covering junior, mid-level, and senior professionals.
→ BuildxGCC – Global Capability Center Solutions Specializing in Build, Operate, Transfer (BOT) model for Global Capability Centers, GCC supports companies in setting up, scaling, and transferring GCCs.
We are looking for a Senior Site Reliability Engineer to manage and improve the reliability, scalability, availability, and performance of production systems across a hybrid/multi-cloud environment .
Own SLIs, SLOs, error budgets , capacity planning, and reliability improvements.
Manage production Kubernetes / Amazon EKS and Red Hat OpenShift environments.
Work across AWS and IBM Cloud , including hybrid connectivity, networking, and disaster recovery.
Build and maintain infrastructure using Terraform and automate operations using Python and Bash .
Design and improve observability using Prometheus, Grafana, OpenTelemetry, Thanos , and logging platforms.
Lead high-severity incident response , postmortems, and corrective actions.
Implement secure and compliant infrastructure practices across cloud environments.
Mentor engineers and contribute to architecture and reliability standards.
4–6 years of experience in SRE / DevOps / Cloud Infrastructure .
Strong hands‑on experience with AWS and Kubernetes .
Experience with EKS, Docker, Helm ; OpenShift is highly preferred.
Strong Terraform, Python, and Bash skills.
Hands‑on experience with Prometheus and Grafana .
Experience with SLIs, SLOs, incident management, and production troubleshooting .
Good understanding of DNS, networking, TLS, VPN, load balancing, and cloud connectivity .
Exposure to regulated environments such as HIPAA, SOC 2, PCI DSS, ISO 27001 , etc.
IBM Cloud, OpenShift, Argo CD/Flux, Istio/Linkerd, Go, AIOps, Chaos Engineering, AWS/IBM certifications, and FinOps experience.