Senior Site Reliability Engineer (SRE) – GCP

Opstree Global

Bengaluru

On-site

INR 3,600,000 - 6,000,000

Full time

6 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Opstree Global is seeking an experienced Senior Site Reliability Engineer (SRE) with 7–8 years of design, implementation, and maintenance of scalable cloud infra on Google Cloud Platform. The role emphasizes reliability, performance, and operational excellence while collaborating with development teams.

The candidate should be strong in GCP, Kubernetes, Terraform, CI/CD, observability, automation, and production operations, with on-call responsibilities as needed.

Qualifications

  • 7–8 years of SRE/DevOps experience.
  • Strong hands-on experience with GCP, Kubernetes, Linux.
  • Proficiency in Infrastructure as Code (Terraform).
  • CI/CD expertise and automation experience.
  • Observability, monitoring, logging, and incident response.

Responsibilities

  • Design, deploy, and manage highly available infrastructure on GCP.
  • Own production reliability, availability, performance, and operational excellence.
  • Build and maintain GKE/Kubernetes clusters and containerized workloads.
  • Implement and manage CI/CD pipelines for apps and infra.
  • Automate operational tasks using Python/Bash and scripts.
  • Develop infrastructure as code using Terraform.
  • Implement monitoring and observability with Cloud Monitoring/Logging and related tools.
  • Define and monitor SLIs, SLOs, SLAs; participate in incident management.

Skills

GCP expertise
Kubernetes & GKE
Linux
Terraform
CI/CD
Observability
Python/Bash scripting
Incident management
Security & cost optimization

Education

Bachelor's degree in Computer Science, Engineering, IT

Tools

GKE
Terraform
GitLab CI/CD
Jenkins
GitHub Actions
ArgoCD
Prometheus/Grafana

Job description

Job Summary

We are looking for an experienced Senior Site Reliability Engineer (SRE) with 7–8 years of experience in designing, implementing, and maintaining highly available, scalable, and reliable cloud infrastructure on Google Cloud Platform (GCP).

The ideal candidate should have strong hands‑on experience with GCP, Kubernetes, Linux, Infrastructure as Code, CI/CD, observability, automation, and production operations. The role will involve improving system reliability, performance, scalability, and operational efficiency while working closely with development and platform teams.

Key Responsibilities
  • Design, deploy, and manage highly available and scalable infrastructure on GCP.
  • Own production infrastructure, reliability, availability, performance, and operational excellence.
  • Build and maintain GKE/Kubernetes clusters and containerized workloads.
  • Implement and manage CI/CD pipelines for application and infrastructure deployments.
  • Automate repetitive operational tasks using Python, Bash, or similar scripting languages.
  • Develop and maintain Infrastructure as Code using Terraform.
  • Implement monitoring, logging, alerting, and observability using Google Cloud Monitoring, Cloud Logging, and other tools.
  • Define and monitor SLIs, SLOs, and SLAs for critical services.
  • Participate in production incident management, troubleshooting, root‑cause analysis, and post‑incident reviews.
  • Perform capacity planning, performance tuning, and scalability assessments.
  • Implement high‑availability, disaster recovery, backup, and business‑continuity strategies.
  • Identify reliability risks and proactively implement solutions to reduce system downtime.
  • Work with development teams to improve application reliability and deployment processes.
  • Establish and improve operational runbooks, automation, and self‑healing mechanisms.
  • Support security and compliance requirements across cloud infrastructure.
  • Participate in on‑call/production support activities when required.
Required Technical Skills
GCP
  • Strong hands‑on experience with Google Cloud Platform.
  • Experience with:
  • GKE
  • Compute Engine
  • VPC
  • IAM
  • Cloud Load Balancing
  • Cloud Storage
  • Cloud SQL
  • Pub/Sub
  • Cloud Monitoring & Logging
  • Secret Manager
  • Good understanding of GCP networking, IAM, security, and cost optimization.
Kubernetes & Containers
  • Strong hands‑on experience with Kubernetes and GKE.
  • Docker/containerization.
  • Kubernetes networking, deployments, services, ingress, configmaps, secrets, RBAC, and troubleshooting.
  • Experience with Helm and Kubernetes operators is preferred.
Infrastructure as Code
  • Strong experience with Terraform.
  • Experience designing reusable Terraform modules and managing infrastructure through Git‑based workflows.
CI/CD & DevOps
  • Strong understanding of CI/CD principles.
  • Hands‑on experience with tools such as:
  • GitLab CI/CD
  • Jenkins
  • GitHub Actions
  • ArgoCD
  • Experience implementing automated deployment and rollback strategies.
Observability & Reliability
  • Experience with monitoring, logging, tracing, and alerting.
  • Hands‑on experience with Prometheus, Grafana, OpenTelemetry or equivalent tools.
  • Strong understanding of:
  • SLIs
  • SLOs
  • SLAs
  • Error Budgets
  • Incident Management
  • RCA
Linux & Networking
  • Strong Linux administration and troubleshooting skills.
  • Good understanding of:
  • TCP/IP
  • DNS
  • HTTP/HTTPS
  • Load Balancing
  • SSL/TLS
  • Network troubleshooting
  • Firewall and security concepts.
Scripting & Automation
  • Strong scripting experience in Python and/or Bash.
  • Ability to build automation tools and operational scripts.
Good to Have
  • GCP Professional Cloud DevOps Engineer certification.
  • Experience with GitOps and ArgoCD.
  • Experience with service mesh such as Istio.
  • Experience with SRE frameworks and practices.
  • Experience with chaos engineering and resilience testing.
  • Experience with distributed systems and microservices architecture.
  • Experience with FinOps/cloud cost optimization.
  • Exposure to security best practices and DevSecOps.
  • Experience working in 24x7 production environments.
Required Soft Skills
  • Strong analytical and problem‑solving skills.
  • Excellent troubleshooting and debugging ability.
  • Good communication and documentation skills.
  • Ability to work independently and take ownership of production systems.
  • Strong incident management and decision‑making skills.
  • Ability to collaborate effectively with Development, Security, Product, and Platform teams.
Education

Bachelor’s degree in Computer Science, Engineering, Information Technology, or a related field is preferred.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE) – GCP Platform
Site Reliability Engineer (SRE) – GCP Platform

ITC Infotech • Bengaluru

On-site
INR 900,000 - 1,300,000
Site Reliability Engineer- GCP
Site Reliability Engineer- GCP

Aziro • Hyderabad

Hybrid
INR 1,400,000 - 2,100,000
Site Reliability Engineer (SRE) - Google Cloud Platform
Site Reliability Engineer (SRE) - Google Cloud Platform

Aziro • Hyderabad

Hybrid
INR 1,500,000 - 3,200,000
Intermediate Applications Developer
Intermediate Applications Developer

UPS • Chennai District

On-site
INR 1,500,000 - 2,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MangoApps • Maharashtra

On-site
INR 4,000,000 - 7,000,000
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

Skillventory • Kamrup Metropolitan

On-site
INR 3,500,000 - 7,000,000
SRE-AWS/GCP
SRE-AWS/GCP

Sailssoftware • Visakhapatnam

On-site
INR 1,500,000 - 2,500,000
Site Reliability Engineer
Site Reliability Engineer

Advance Career Solutions • Pune District

On-site
INR 1,200,000 - 1,800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MangoApps INC. • Pune District

On-site
INR 1,400,000 - 1,800,000
Site Reliability Engineer - Google Cloud Platform
Site Reliability Engineer - Google Cloud Platform

GAMMASTACK • Kolkata District

On-site
INR 1,800,000 - 2,800,000