Enterprise Observability Platform Engineer

Be a Catalyst

Gurugram District

On-site

INR 1,500,000 - 2,000,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Be a Catalyst is seeking an experienced Consultant – Enterprise Observability Platform Engineer to design and develop an enterprise-wide observability platform. This role involves driving platform architecture and establishing observability standards.

The ideal candidate will have over 5 years of experience in Site Reliability Engineering or DevOps, and hands-on experience with tools like Grafana and Prometheus, aimed at supporting cloud-native and hybrid environments.

Qualifications

  • 5+ years of experience in Site Reliability Engineering, Platform Engineering, or DevOps.
  • Strong understanding of Observability concepts including Logs, Metrics, and Traces.
  • Hands-on experience with observability tools like Grafana and Prometheus.

Responsibilities

  • Define and evolve the Enterprise Observability Platform architecture.
  • Design reference architectures and observability patterns for applications.
  • Evaluate and integrate observability tools into the platform.
  • Automate provisioning and onboarding for the observability services.
  • Collaborate with various teams to drive observability adoption.

Skills

Site Reliability Engineering (SRE)
Observability concepts
Grafana
Prometheus
Elastic
Splunk
Datadog
OpenTelemetry
Scripting with Python
Kubernetes

Job description

We are looking for an experienced Consultant – Enterprise Observability Platform Engineer to design, develop, and govern an enterprise-wide observability platform. The ideal candidate will drive platform architecture, establish observability standards, enable developer self‑service, and support engineering resilience across cloud‑native, hybrid, and legacy environments.

Responsibilities
  • Define and evolve the Enterprise Observability Platform architecture integrating logs, metrics, traces, events, and alerts
  • Design reference architectures and observability patterns for cloud‑native and monolithic applications
  • Evaluate and integrate observability tools including OpenTelemetry, Prometheus, Grafana, Elastic, Splunk, Datadog, and New Relic
  • Define enterprise observability standards, SLOs, SLIs, instrumentation guidelines, and telemetry governance
  • Build and maintain scalable, highly available, and cost‑optimized observability platform services
  • Automate provisioning, onboarding, alert configuration, and tenant lifecycle management
  • Enable self‑service capabilities through instrumentation kits, dashboards, alert templates, and troubleshooting guides
  • Collaborate with Platform Engineering, Developer Experience, SRE, and Application teams to drive observability adoption
  • Lead application onboarding, migration, and telemetry integration initiatives
  • Define observability KPIs, telemetry baselines, and portfolio‑level monitoring strategies
Qualifications
  • 5+ years of experience in Site Reliability Engineering (SRE), Platform Engineering, or DevOps
  • Strong understanding of Observability concepts – Logs, Metrics, Traces, Events, SLOs, SLIs, RED & USE models
  • Hands‑on experience with Grafana, Prometheus, Elastic, Splunk, Datadog, OpenTelemetry, or New Relic
  • Strong scripting and automation skills using Python, Go, Bash, or Terraform
  • Experience with Kubernetes, Container Orchestration, AWS, and Azure
  • Experience building or supporting Enterprise Observability Platforms
  • Knowledge of Multi‑Tenant Observability Systems and Governance‑as‑Code
  • Experience with CI/CD integration and developer enablement
  • Strong troubleshooting, automation, and platform engineering mindset
  • Excellent communication and stakeholder management skills
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Observability & Sr. Observability Platform Engineer
Observability & Sr. Observability Platform Engineer

American Express Global Business Travel • Bengaluru

On-site
INR 3,500,000 - 7,500,000
SRE Observability Engineer
SRE Observability Engineer

Awign • Hyderabad

On-site
INR 4,200,000 - 6,500,000
Site Reliability Engineer
Site Reliability Engineer

TerraGiG • India

On-site
INR 1,500,000 - 2,500,000
Monitoring Specialist - Observability Platform Engineer
Monitoring Specialist - Observability Platform Engineer

Festo India • Bengaluru

On-site
INR 2,500,000 - 3,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Brillio • Bengaluru Urban

On-site
INR 1,200,000 - 2,000,000
Senior Observability Architect (Prometheus and Grafana)
Senior Observability Architect (Prometheus and Grafana)

Keka Technologies Private Limited • Bengaluru

On-site
INR 3,600,000 - 7,200,000
Senior Software Engineer
Senior Software Engineer

NVIDIA • Bengaluru

On-site
INR 4,000,000 - 7,000,000
SRE
SRE

VXI Global Solutions • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Senior Software Engineer
Senior Software Engineer

NVIDIA Corporation • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Site Reliability Engineer (SRE) / Observability Engineer
Site Reliability Engineer (SRE) / Observability Engineer

N Human Resources & Management Systems • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Hybrid work
Certification reimbursement
Structured learning