Observability Engineer (SRE)

Cloudstepin

Hyderabad

On-site

INR 1,200,000 - 1,800,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A dynamic tech company in Hyderabad is seeking an Observability Engineer (SRE) to build and manage monitoring systems. The role requires a detail-oriented approach and collaboration with teams to improve platform reliability. Candidates should have 2+ years of experience in SRE or DevOps, strong skills in observability tools like Prometheus and Grafana, and proficiency in scripting and infrastructure as code. This is a full-time position with opportunities for professional growth.

Qualifications

  • 2+ years of experience in SRE, DevOps, or Infrastructure Engineering.
  • Strong experience with observability tools.
  • Knowledge of distributed systems and cloud platforms (AWS/GCP/Azure).
  • Proficient in scripting (Python, Bash, etc.) and infrastructure as code (Terraform, Ansible).
  • Familiarity with CI/CD pipelines and container orchestration (Docker, Kubernetes).
  • Strong understanding of system performance, monitoring and reliability principles.

Responsibilities

  • Design and implement observability solutions (metrics, logs, traces).
  • Build dashboards and alerts using Prometheus, Grafana, or ELK.
  • Collaborate to improve system reliability and reduce MTTR.
  • Automate monitoring for distributed services.
  • Implement SLOs/SLIs and improve incident response.
  • Optimize logging and tracing for performance.
  • Lead root cause analysis and post-incident reviews.

Skills

Observability tools
Scripting (Python, Bash)
Infrastructure as code (Terraform, Ansible)
System performance monitoring
CI/CD pipelines

Education

Any degree

Tools

Prometheus
Grafana
ELK
DataDog
New Relic
Terraform
Ansible
Docker
Kubernetes
Docker
Kubernetes

Job description

Job Title:Observability Engineer (SRE)

Job Category:Technical

Job Type:Full Time

City:Hyderabad

Bond:No

Experience:2+

Qualification:Any degree

Salary:Depends on Skill

About the Role:

We’re looking for a passionate and detail-oriented Observability Engineer (SRE) to build and manage monitoring, alerting and telemetry systems that keep our platforms reliable, fast and scalable.

You’ll work at the intersection of software engineering and systems operations—ensuring uptime, improving performance and empowering teams with visibility into production systems.

Key Responsibilities:
  • Design and implement end-to-end observability solutions (metrics, logs, traces)
  • Build dashboards, alerts and insights using tools like Prometheus, Grafana, ELK, DataDog or New Relic
  • Collaborate with SREs and developers to improve system reliability and reduce MTTR
  • Automate monitoring and alerting for distributed services and cloud infrastructure
  • Implement SLOs/SLIs and improve incident response workflows
  • Optimize logging and tracing pipelines for performance and cost
  • Lead root cause analysis and drive post-incident reviews
Requirements:
  • 2+ years in SRE, DevOps or Infrastructure Engineering
  • Strong experience with observability tools: Prometheus, Grafana, ELK, Datadog, New Relic, etc.
  • Knowledge of distributed systems and cloud platforms (AWS/GCP/Azure)
  • Proficient in scripting (Python, Bash, etc.) and infrastructure as code (Terraform, Ansible)
  • Familiarity with CI/CD pipelines and container orchestration (Docker, Kubernetes)
  • Strong understanding of system performance, monitoring and reliability principles

Job Category:Cloud Engineer

Job Type:Full Time

Job Location:Hyderabad

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE) / Observability Engineer
Site Reliability Engineer (SRE) / Observability Engineer

N Human Resources & Management Systems • Hyderabad

Hybrid
INR 4,000,000 - 7,000,000
Hybrid work
Certification reimbursement
Structured learning
Observability / SRE Engineer
Observability / SRE Engineer

Synapse Business Systems • Hyderabad, Bengaluru

On-site
INR 1,500,000 - 3,000,000
SRE - Site Reliability Engineering
SRE - Site Reliability Engineering

Build & Hire • Pune District

On-site
INR 1,500,000 - 2,300,000
SRE Observability Engineer
SRE Observability Engineer

Awign • Hyderabad

On-site
INR 4,200,000 - 6,500,000
Observability Engineer
Observability Engineer

Weekday (YC W21) • Hyderabad

On-site
INR 4,200,000 - 7,000,000
Observability Engineer
Observability Engineer

Weekday (YC W21) • Chennai District

On-site
INR 3,500,000 - 7,000,000
Observability Engineer
Observability Engineer

Weekday (YC W21) • Mumbai

On-site
INR 4,000,000 - 6,000,000
Observability Engineer
Observability Engineer

Weekday (YC W21) • Bengaluru

On-site
INR 3,500,000 - 6,000,000
SRE New Relic
SRE New Relic

Tekskills • Hyderabad, Chennai District, Bengaluru

Hybrid
INR 400,000 - 700,000
Observability Engineer
Observability Engineer

ITC Infotech • Gurugram District

On-site
INR 4,000,000 - 7,000,000