Observability Expert Grafana

Nestlé

Bengaluru South

Hybrid

INR 1,200,000 - 2,400,000

Full time

23 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Nestlé Nespresso is seeking an Observability Platform Expert to lead hands-on implementation, configuration, and continuous improvement of enterprise observability across critical services, infrastructure, and customer journeys. You will monitor, log, trace, and correlate telemetry to improve end-to-end visibility in hybrid/multi-cloud environments.

You will onboard services, validate instrumentation, develop dashboards and alerts, troubleshoot telemetry pipelines, and support production

Qualifications

  • Bachelor's degree or equivalent professional experience.
  • Minimum five years of experience in observability, monitoring, SRE or production operations.
  • Strong hands-on experience implementing and operating Grafana in complex enterprise environments.

Responsibilities

  • Maintain, run, and continuously improve our monitoring infrastructure.
  • Setup new monitoring capabilities for existing infrastructure and applications (dashboards, alerts, metrics).
  • Ensure stability of mission-critical systems by setting up alerts and auto-remediation.
  • Setup event-driven automation based on monitoring events.
  • Collaborate with international teams to onboard services and validate instrumentation.

Skills

Grafana dashboards
PromQL
OpenTelemetry
Kubernetes
Linux
Cloud platforms

Education

Bachelor's degree in Computer Science or related

Tools

Grafana
Prometheus
Loki
Tempo
Mimir
OpenTelemetry
Kubernetes

Job description

Job Description:
OBSERVABILITY EXPERT (GRAFANA)

We are looking for an Observability Expert to be part of our Nestlé Nespresso Digital and Tech Team. At Nespresso, our Digital & Tech teams are at the heart of our innovation journey, a space where we continue to invest, evolve, and grow.

Position Snapshot:
  • Location: Bengaluru, Karnataka, India
  • Type of Contract: Permanent
  • Grade: Band 2
  • Type of work: Hybrid
  • Work Language: Fluent Business English
The Role:

As a Observability Platform Expert at Nespresso, you will be responsible for the hands‑on implementation, configuration, operation, and continuous improvement of enterprise observability capabilities across our critical digital services, applications, infrastructure, networks, and customer journeys.

You will use your deep expertise in monitoring, logging, distributed tracing, Real User Monitoring, and telemetry correlation to improve End‑to‑End visibility across complex hybrid and multi‑cloud environments. You will help teams detect service degradation earlier, investigate incidents faster, identify root causes, and reduce customer and business impact.

Working closely with application owners, operations, platform teams, Enterprise Architecture, security, and external partners, you will onboard services into the observability platform, validate instrumentation coverage, develop dashboards and actionable alerts, troubleshoot telemetry pipelines, and support production incident investigations.

The role requires a strong hands‑on mindset. You will implement solutions directly, lead technical evaluations and Proofs of Concept, and translate business and operational requirements into measurable observability outcomes.

Knowledge of observability architecture is a strong advantage. You may contribute to target architecture, platform integration patterns, telemetry standards, application onboarding models, security and data‑governance requirements, scalability, resilience, and cost optimization. You will also support architecture reviews and help ensure solutions are aligned with OpenTelemetry and other relevant industry standards.

You will stay current with emerging observability practices and technologies, including AIOps, AI-assisted investigation, automation, and Agentic AI, evaluating them based on practical value, operational readiness, governance, and cost.

In This Role, You Will:
  • Maintain, run, and continuously improve our monitoring infrastructure.
  • Setup new monitoring capabilities for the existing/new infrastructure and applications (dashboards, alerts, metrics…)
  • Ensure the stability of the mission‑critical systems by setting up alerts, notification, and auto‑remediation.
  • Setup event‑driven automation based on events from monitoring systems.
  • Collaborate with an international, self‑responsible and agile team who is working with the newest toolset.
What We’re Looking For:
  • Bachelor’s degree in Computer Science, Systems Engineering or a related discipline, or equivalent professional experience.
  • Minimum five years of experience in observability, monitoring, Site Reliability Engineering or production operations.
  • Strong hands‑on experience implementing and operating Grafana in complex enterprise environments.
Advanced experience with:
  • Grafana dashboards, variables, transformations and data‑source configuration.
  • Prometheus and PromQL.
  • Grafana Mimir or managed Prometheus environments.
  • Grafana Loki and LogQL.
  • Grafana Tempo and distributed tracing.
  • Grafana Alloy for telemetry collection and forwarding.
  • Grafana Alerting, contact points, notification policies and alert routing.
  • Experience designing and troubleshooting dashboards and alerts for applications, infrastructure, databases, networks and business services.
  • Strong knowledge of OpenTelemetry, including:
  • OpenTelemetry Collector architecture.
  • OTLP ingestion and export.
  • Metrics, logs and traces.
  • Context propagation and trace correlation.
  • Sampling and telemetry‑processing strategies.li>
  • Experience onboarding applications and infrastructure into the Grafana observability ecosystem.
  • Experience correlating metrics, logs and traces to support End‑to‑End troubleshooting and root‑cause analysis.
  • Experience troubleshooting production incidents across hybrid and multicloud environments.
  • Working knowledge of Linux, containers and Kubernetes.
  • Experience with at least one major cloud platform such as Azure, AWS, OCI or GCP.
  • Experience integrating Grafana with ITSM, incident‑management and DevOps tools such as ServiceNow, Jira, PagerDuty or CI/CD platforms.
  • Proven experience leading technical evaluations, proofs of concept and production implementations.
  • Strong English communication skills and experience working with global, distributed teams.
Extra Skills That Set You Apart:
  • Experience having worked in a global environment and with virtual teams.
  • Solid understanding and hands on experience using Ansible and Kubernetes.
  • Experience working with Ecommerce platforms.
  • Good understanding of MS Azure.
  • Experience with UNIX/Linux and UNIX shell scripting knowledge (Bash and Python).
  • Architecture experience — strong advantage
  • Experience designing enterprise observability architectures for hybrid, multicloud and on‑premises environments.
  • Ability to define telemetry collection,
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Grafana Developer
Grafana Developer

Ericsson GmbH • Bengaluru

On-site
INR 1,800,000 - 3,600,000
Grafana Developer Bangalore,Karnataka,India Service Delivery Posted 6 hours ago
Grafana Developer Bangalore,Karnataka,India Service Delivery Posted 6 hours ago

Ericsson GmbH • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Observability Engineer - Grafana - Long term contract
Observability Engineer - Grafana - Long term contract

Cygnus Professionals Inc. • Mumbai

Hybrid
INR 1,800,000 - 2,400,000
Associate Observability Architect | PST | Remote
Associate Observability Architect | PST | Remote

Embedded Shishya • Gopalganj

Hybrid
INR 13,263,000 - 15,935,000
RSUs
Remote-first culture
Senior Observability Architect (Prometheus and Grafana)
Senior Observability Architect (Prometheus and Grafana)

Keka Technologies Private Limited • Bengaluru

On-site
INR 3,600,000 - 7,200,000
Site Reliability Engineer (SRE) / Observability Engineer
Site Reliability Engineer (SRE) / Observability Engineer

N Human Resources & Management Systems • Hyderabad

Hybrid
INR 4,000,000 - 7,000,000
Hybrid work
Certification reimbursement
Structured learning
Software Engineer
Software Engineer

Saint-Gobain Group in India • Chennai District

On-site
INR 900,000 - 1,200,000
SRE Observability Engineer
SRE Observability Engineer

Awign • Hyderabad

On-site
INR 4,200,000 - 6,500,000
Advanced Engineer Software - AI Ops Fullstack [T500-26632]
Advanced Engineer Software - AI Ops Fullstack [T500-26632]

Albertsons Companies India • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Observability Engineer
Observability Engineer

Coforge • Hyderabad, Pune District, Greater Noida

Hybrid
INR 2,400,000 - 4,200,000