SRE (Site Reliability Engineer)

VXI

India

On-site

INR 1,400,000 - 2,200,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

VXI in India is seeking a skilled SRE Engineer to design, implement, and manage observability across our cloud infrastructure and applications. You will work hands-on with Prometheus, Grafana, Cloud Monitoring, and OpenTelemetry to instrument distributed systems and monitor health and performance.

The role involves building dashboards and alerts, integrating with Google Cloud Platform and SolarWinds, correlating logs, metrics, and traces to reduce MTTR, and collaborating with SRE, DevOps, and

Qualifications

  • Hands-on with Prometheus and Grafana for monitoring and alerting.
  • Proficiency with OpenTelemetry for instrumenting distributed systems.
  • Working knowledge of observability tools in Google Cloud (Cloud Monitoring, Logging, Trace).
  • Exposure to SolarWinds for network and infrastructure monitoring.
  • Solid understanding of metrics, logs, and traces and ability to correlate telemetry data.

Responsibilities

  • Design and maintain observability pipelines using OpenTelemetry, Prometheus, and Grafana.
  • Build dashboards and alerts to monitor system health, application performance, and business KPIs.
  • Integrate observability solutions with Google Cloud Platform services and SolarWinds.
  • Correlate logs, metrics, and traces to troubleshoot incidents and reduce MTTR.
  • Collaborate with SREs, DevOps, and development teams to improve end-to-end system observability.
  • Implement best practices for telemetry data collection, enrichment, storage, and visualization.

Skills

Prometheus
Grafana
OpenTelemetry
Cloud Monitoring
SolarWinds
Python
Bash
Infrastructure as Code
Telemetry data types

Job description

It's fun to work in a company where people truly BELIEVE in what they are doing! We're committed to bringing passion and customer focus to the business.

Job Summary:

We are seeking a skilled SRE Engineer to design, implement, and manage robust observability solutions across our cloud infrastructure and applications. The ideal candidate will have hands-on experience with Prometheus, Grafana, Cloud Monitoring, and OpenTelemetry, along with exposure to SolarWinds. You should be comfortable working with metrics, logs, and traces, and be able to correlate telemetry data to proactively detect, diagnose, and resolve performance issues.

Key Responsibilities:
  • Design and maintain observability pipelines using OpenTelemetry, Prometheus, and Grafana.
  • Build dashboards and alerts to monitor system health, application performance, and business KPIs.
  • Integrate observability solutions with Google Cloud Platform services and SolarWinds.
  • Correlate logs, metrics, and traces to troubleshoot incidents and reduce MTTR.
  • Collaborate with SREs, DevOps, and development teams to improve end-to-end system observability.
  • Implement best practices for telemetry data collection, enrichment, storage, and visualization.
Requirements:
  • Strong experience with Prometheus and Grafana for monitoring and alerting.
  • Proficiency in OpenTelemetry for instrumenting distributed systems.
  • Working knowledge of observability tools in Google Cloud (e.g., Cloud Monitoring, Logging, Trace).
  • Exposure to SolarWinds for network and infrastructure monitoring.
  • Solid understanding of telemetry data types: metrics, logs, and traces.
  • Ability to correlate and analyze multi-source observability data.
  • Scripting skills (Python, Bash) and familiarity with Infrastructure-as-Code is a plus.
Preferred Qualifications:
  • Experience in Site Reliability Engineering or Platform Engineering roles.
  • Knowledge of SLIs/SLOs and performance benchmarking.
  • Experience with APM tools (e.g., Datadog, New Relic) is a plus.

If you like wild growth and working with happy, enthusiastic over-achievers, you'll enjoy your career with us!

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE
SRE

VXI Global Solutions • Hyderabad

On-site
INR 1,200,000 - 1,800,000
SRE
SRE

Tekion • Bengaluru

On-site
INR 1,200,000 - 1,800,000
SRE Observability Engineer
SRE Observability Engineer

Awign • Hyderabad

On-site
INR 4,200,000 - 6,500,000
Site Reliability Engineer (SRE) – DevOps Infrastructure
Site Reliability Engineer (SRE) – DevOps Infrastructure

PQAngels Technologies Pvt. Ltd. • Bengaluru

On-site
INR 1,200,000 - 2,100,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

C1X • Chennai District

On-site
INR 1,800,000 - 3,200,000
Site Reliability Engineer
Site Reliability Engineer

InOpTra Digital • Bengaluru

On-site
INR 1,200,000 - 2,000,000
Site Reliability Engineer
Site Reliability Engineer

Peoplefy • Pune District

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer
Site Reliability Engineer

Snapmint • Gurugram District

On-site
INR 800,000 - 1,200,000
SRE Lead
SRE Lead

Acldigital • Ahmedabad District

On-site
INR 1,500,000 - 2,000,000
Site Reliability Engineer (SRE) / Observability Engineer
Site Reliability Engineer (SRE) / Observability Engineer

N Human Resources & Management Systems • Hyderabad

On-site
INR 4,000,000 - 7,000,000
Hybrid work
Certification reimbursement
Structured learning