Site Reliability Engineer

Indihire Consultants

Hyderabad

On-site

INR 1,500,000 - 2,800,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Indihire Consultants in Hyderabad invites an experienced Site Reliability Engineer to design, build, and operate scalable cloud-based applications and infrastructure. You will be embedded with the Development team while reporting to the infrastructure leader, focusing on reliability, performance, and deploy practices.

The role demands hands-on experience with GCP or similar cloud platforms, containers (Docker/Kubernetes), and IaC tools like Terraform.

Qualifications

  • 5+ years designing, developing, delivering, and operating scalable applications and infrastructure.
  • Bachelor's or advanced degree in Computer Science, Information Systems, or a related field
  • Google Professional Cloud Architect or similar certification desired
  • Advanced hands-on experience with CI/CD methodologies and technologies
  • Advanced experience with cloud environments (Google Cloud Platform or similar)
  • Strong experience with containers (Docker) and IaC (Terraform)
  • Working knowledge of Helm and service mesh concepts like Istio
  • Proficiency with scripting (Shell, Python) in Linux
  • Ability to work independently and in cross-functional teams
  • Excellent understanding of internet concepts and protocols

Responsibilities

  • Maintain availability, performance, and scalability of critical services and production environments.
  • Collaborate with developers to design reliable applications and improve deployment practices.
  • Break down walls and build trust between developers and infrastructure teams.
  • Own end-to-end CI/CD including automated testing, observability, dependency management, and operational concerns.
  • Improve CI/CD pipeline reliability, traceability, and security.
  • Build automation for provisioning, configuration, deployments, and incident response.
  • Improve observability using metrics, logs, distributed tracing, dashboards, and alerting.
  • Participate in on-call rotation, lead incident response, and drive root cause analysis.
  • Conduct capacity planning, chaos testing, and reliability reviews.
  • Implement infrastructure-as-code using Terraform, Helm, Jenkins/GitHub Actions, etc.
  • Optimize CI/CD pipelines and ensure safe, repeatable deployments (ArgoCD).
  • Champion SRE principles: SLIs/SLOs, error budgets, toil reduction, postmortems.
  • Embrace a culture of enablement, customer service, continuous improvement, transparency, and fiscal responsibility.
  • Perform other duties as directed

Skills

Scalability design
CI/CD pipelines
Docker
Kubernetes
Terraform
Helm
Istio
Observability
Shell scripting
Python
GCP / cloud platforms

Education

Bachelor's in CS/Information Systems

Tools

Terraform
Helm
Jenkins/GitHub Actions
Istio
Docker
Kubernetes
CI/CD tooling

Job description

Role & responsibilities
  • Maintain availability, performance, and scalability of critical services and production environments.
  • Collaborate closely with developers to design reliable applications and improve deployment practices (you will be embedded in the Development team but reporting to infrastructure leader).
  • Break down walls and build trust between developers and infrastructure teams
  • Participate in end-to-end application ownership throughout the CI/CD process, including automated testing, observability, dependency management, and other operational concerns.
  • Improve CI/CD pipeline reliability, traceability, and security
  • Build automation for provisioning, configuration, deployments, and incident response.
  • Improve observability using metrics, logs, distributed tracing, dashboards, and alerting.
  • Participate in on-call rotation, lead incident response, and drive root cause analysis.
  • Conduct capacity planning, chaos testing, and reliability reviews.
  • Implement infrastructure-as-code using Terraform, Helm, Jenkins/GitHub Actions/etc.
  • Optimize CI/CD pipelines and ensure safe, repeatable deployments (i.e. ArgoCD).
  • Champion SRE principles: SLIs/SLOs, error budgets, toil reduction, problem management, blameless postmortems.
  • Embrace a culture of enablement, customer service, continuous improvement, transparency, and fiscal responsibility
  • Perform other duties as directed
Preferred candidate profile

KNOWLEDGE, SKILLS AND ABILITIES
  • 5+ years designing, developing, delivering, and operating scalable, available, high-performance applications (Java and node.js, etc.) and infrastructure
  • Bachelor's or advanced degree in Computer Science, Information Systems, or a related field
  • Familiarity with modern application languages and concepts, with hands‑on e-commerce software development experience preferred
  • Google Professional Cloud Architect or similar certification desired
  • Advanced hands‑on experience with continuous integration and delivery / deployment methodologies and technologies
  • Advanced experience with computer, networking, security, storage, monitoring, logging, database, and other technologies in Google Cloud Platform or similar major cloud environment
  • Strong experience with containerization (e.g. Docker), Kubernetes, and Infrastructure as Code (Terraform preferred)
  • Working knowledge of Helm and Service Mesh (e.g. Istio)
  • Proficient understanding of microservices principles and orchestration
  • Excellence in navigating and prioritizing multiple simultaneous responsibilities of varying scope and complexity
  • Ability to effectively articulate technical concepts to audiences at all organizational levels via oral, written, and other non‑verbal communications
  • Demonstrated desire and ability to be self‑directed, take ownership of issues, and establish a prominent level of credibility
  • Ability to work well independently and within dynamic, cross-functional teams
  • Excellent understanding of Internet concepts, technologies and protocols (TCP/IP, DNS, HTTP, TLS / SSL, etc.)
  • Experience with rapid detection and resolution of technical issues using various monitoring and application performance management tools
  • Proficiency with shell scripting, Python and/or other scripting languages in a Linux environment
  • Ability to operate effectively under pressure, both independently and in collaboration with other resources
  • Ability to rapidly learn new technologies via mentoring, formal training, independent research and testing
  • A genuine desire and willingness to share knowledge effectively with others
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

MangoApps • Maharashtra

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

MangoApps INC. • Pune District

On-site
INR 1,400,000 - 1,800,000
Senior Platform Engineer
Senior Platform Engineer

Quantiphi Analytics Solutions • Mumbai, Bengaluru

Hybrid
INR 1,200,000 - 2,400,000
Site Reliability Engineering (SRE) Lead
Site Reliability Engineering (SRE) Lead

SID Global Solutions • Hyderabad

On-site
INR 3,000,000 - 5,000,000
Senior Infrastructure Engineer
Senior Infrastructure Engineer

Equiti Group • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

NCR Voyix • Chennai District

On-site
INR 3,000,000 - 5,400,000
GCP Cloud Engineer
GCP Cloud Engineer

ITC Infotech • Bengaluru

On-site
INR 1,800,000 - 2,400,000
Lead DevOps Engineer
Lead DevOps Engineer

WizCommerce • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Senior Team Lead | Engineering, AI & Data - Engineering | Site Reliability Engineering
Senior Team Lead | Engineering, AI & Data - Engineering | Site Reliability Engineering

Deloitte & Touche GmbH Wirtschaftsprüfungsgesellschaft • Bengaluru

On-site
INR 2,000,000 - 3,000,000
Site Reliability Engineer (SRE) – Core IT Infrastructure
Site Reliability Engineer (SRE) – Core IT Infrastructure

TECEZE • Chennai District

On-site
INR 1,000,000 - 2,000,000