AWS Observability Engineer

Tata Consultancy Services

New Delhi, Dadri

On-site

INR 1,500,000 - 2,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tata Consultancy Services is seeking an Observability Engineer in New Delhi to design and operate observability solutions for cloud-native platforms. The role involves maintaining metrics, logs, and traces while ensuring system reliability and performance. Ideal candidates should possess a strong understanding of observability concepts, hands-on experience with tools like OpenTelemetry, Prometheus, and Grafana, and 6 to 10 years of relevant experience. Join a collaborative team to enhance system monitoring and operational excellence.

Qualifications

  • Strong understanding of observability concepts (metrics, logs, traces).
  • Hands-on experience with observability tools and cloud services.
  • 6 to 10 years of experience in a relevant field.

Responsibilities

  • Design, implement, and maintain observability platforms.
  • Create and maintain SLOs, SLIs, and alerting strategies.
  • Support production incident investigations and perform root cause analysis.

Skills

OpenTelemetry
Prometheus
Grafana
Kubernetes
AWS
Terraform
CI/CD pipelines
Metrics
Logs
Traces

Tools

Helm
GitOps
Linux

Job description

Job Title: Observability Engineer

Role Overview

We are looking for a highly motivated Observability Engineer to design, implement, and operate end to end observability solutions for modern, cloud native platforms.

The role focuses on building and maintaining metrics, logs, and tracing (MELT) pipelines using industry standard tools and ensuring high system reliability, performance, and visibility.

You will work closely with SRE, DevOps, Platform, and Application teams to improve system monitoring, troubleshoot production issues, and drive a culture of operational excellence.

________________________________________

Key Responsibilities
Observability Platform Engineering
  • Design, implement, and maintain observability platforms using OpenTelemetry, Prometheus, Grafana, Loki, and Tempo
  • Build scalable pipelines for metrics, logs, and distributed traces
  • Define and enforce observability standards across teams
Monitoring & Alerting
  • Create and maintain SLOs, SLIs, and alerting strategies
  • Design actionable alerts that reduce noise and prevent alert fatigue
  • Configure dashboards, alerts, and runbooks for production systems
Kubernetes & Cloud Observability
  • Implement observability for Kubernetes (EKS/GKE/AKS) workloads
  • Enable pod level, node level, and cluster level visibility
  • Integrate observability with cloud services (AWS/GCP/Azure)
Incident Response & Troubleshooting
  • Support production incident investigations using logs, metrics, and traces
  • Perform root cause analysis (RCA) and post incident reviews
  • Improve MTTR by enhancing observability coverage
Automation & Optimization
  • Automate observability deployment using Helm, Terraform, or GitOps
  • Optimize cost and performance of telemetry pipelines
  • Improve data retention, sampling, and aggregation strategies
Collaboration & Enablement
  • Partner with development teams to onboard applications to observability
  • Provide guidance on instrumentation best practices
  • Document observability architectures and operational playbooks
Required Skills
Core Technical Skills
  • Strong understanding of observability concepts (metrics, logs, traces)
  • Hands on experience with OpenTelemetry (SDKs, Agents, Gateways), Prometheus (scraping, recording rules, alerts), Grafana (dashboards, alerts, correlations), Loki or other log aggregation systems, Tempo/Jaeger for distributed tracing
Cloud & Platform
  • Experience with Kubernetes
  • Experience running workloads on AWS (preferred) or other clouds
  • Familiarity with cloud services (EKS, EC2, IAM, S3, Load Balancers)
DevOps & SRE Tooling
  • CI/CD pipelines (GitHub Actions, Jenkins, GitLab)
  • Infrastructure as Code (Terraform / CloudFormation)
  • Linux and networking fundamentals
Experience Level
  • 6 to 10 Years
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Observability Engineer
Observability Engineer

Weekday (YC W21) • Mumbai

On-site
INR 4,000,000 - 6,000,000
Observability Engineer
Observability Engineer

Weekday (YC W21) • Bengaluru

On-site
INR 3,500,000 - 6,000,000
Observability Engineer
Observability Engineer

Weekday (YC W21) • Chennai District

On-site
INR 3,500,000 - 7,000,000
Observability Engineer
Observability Engineer

Weekday (YC W21) • Hyderabad

On-site
INR 4,200,000 - 7,000,000
Enterprise Observability Platform Engineer
Enterprise Observability Platform Engineer

Be a Catalyst • Gurugram District

On-site
INR 1,500,000 - 2,000,000
SRE Observability Engineer
SRE Observability Engineer

Awign • Hyderabad

On-site
INR 4,200,000 - 6,500,000
Observability Engineer - Grafana & Prometheus
Observability Engineer - Grafana & Prometheus

Zensar • Bengaluru

On-site
INR 1,800,000 - 3,000,000
Senior Engineer Observability Platform Engineering
Senior Engineer Observability Platform Engineering

Visara Human Capital Services • Hyderabad, Bengaluru

On-site
INR 900,000 - 1,500,000
Senior Observability Engineer - Grafana & Prometheus
Senior Observability Engineer - Grafana & Prometheus

Zensar Technologies • Bengaluru

On-site
INR 2,800,000 - 4,200,000
Sony - Lead Platform Engineer - Observability Services
Sony - Lead Platform Engineer - Observability Services

Sony India Software Centre • Bengaluru

On-site
INR 1,200,000 - 1,500,000