Senior Site Reliability Engineer (Noida, BLR, India)

Level AI

India

On-site

INR 2,500,000 - 4,500,000

Full time

9 days ago
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Level AI is seeking a Senior SRE at the intersection of backend engineering, infrastructure operations, and FinOps. The role focuses on reducing Kubernetes overprovisioning, driving right-sizing, and maintaining cost telemetry to guide backend teams.

Responsibilities include building tooling, dashboards, and processes that empower backend groups to own cost and reliability budgets, and ensuring robust surface area coverage for cost-at-scale and reliability.

Qualifications

  • 4-5 years of hands-on systems experience
  • Backend engineering depth with production Python/Go/Rust
  • Kubernetes at scale and cost-aware autoscaling
  • Cloud and on-prem infrastructure with GCP and Terraform
  • GPU workloads and throughput profiling
  • Observability with metrics/traces/logs and SLOs
  • FinOps mindset to convert infrastructure choices into cost outcomes
  • Security baseline for platform changes

Responsibilities

  • Drive infrastructure cost efficiency and FinOps
  • Run GPU throughput optimization on on-premise clusters
  • Build tooling and dashboards for cost and reliability budgets
  • Provide reliability instrumentation and ownership of surface area
  • Take on defined security workstreams to support platform changes

Skills

Python
Go/Rust
Backend engineering
Observability
FinOps
Security baseline

Tools

Kubernetes
GCP
Terraform
CI/CD
Cast AI
Karpenter

Job description

About Level AI

Level AI is on a mission to turn every customer interaction into a strategic advantage. Our AI-native platform helps enterprises transform contact centers from cost centers into engines of customer intelligence, operational efficiency, and business growth. By combining advanced AI with deep domain understanding of customer experience, Level AI empowers teams to unlock actionable insights, automate workflows, and deliver more consistent, higher-quality support across the customer journey.

About Level AI

Level AI is on a mission to turn every customer interaction into a strategic advantage. Our AI-native platform helps enterprises transform contact centers from cost centers into engines of customer intelligence, operational efficiency, and business growth. By combining advanced AI with deep domain understanding of customer experience, Level AI empowers teams to unlock actionable insights, automate workflows, and deliver more consistent, higher-quality support across the customer journey.

Headquartered in Mountain View, California, Level AI is a Series C company backed by leading investors including Battery Ventures and ENIAC. Our platform leverages Large Language Models and Custom Small Language Models (SLMs) to power AI Agents across the entire CX journey—customer-facing agents, agent‑assist, and backend automation—along with deep conversation analytics for QA, coaching, and insights.

About The Role

The Senior SRE will be positioned at the intersection of backend engineering, infrastructure operations, and FinOps. The role is explicitly broader than a traditional DevOps engineer and explicitly more hands‑on than a pure architect.

What You'll Be Liable For
  • Infrastructure cost efficiency and FinOps. Own the continued reduction of Kubernetes overprovisioning, drive right-sizing programs, and maintain the cost telemetry that backend teams use to make decisions.
  • GPU throughput optimization. Run a structured experimentation program on on‑premise GPU clusters, partnering with AI service owners. Lead by the Engineering leadership, with this role providing the experimental bandwidth.
  • Backend enablement, not ownership absorption. Build the tooling, dashboards, and processes that let backend teams from other groups own their own cost and reliability budgets. The deliverable is leverage, not headcount‑shaped work.
  • Reliability instrumentation. As the infra team owns most of the instrumentation across new and offline flows, this role takes a central seat in making sure that surface area is captured properly for both cost‑at‑scale and reliability.
  • Selective security workstreams. Take on a defined slice of the active security work so that senior DevOps engineers are not the single point of execution for security‑adjacent platform changes.
We'll love to explore more about you if you have:
  • This role explicitly requires 4-5 years of hands‑on systems experience. We are not looking for someone who will lean entirely on AI tooling to discover what to do; we are looking for someone who already knows what to ask, and can use AI tooling as a force multiplier on top of that judgement.

Backend engineering depth: production experience in Python, Go/Rust, comfortable owning services end to end, able to read and reason about backend code across teams.

Kubernetes at scale: scheduler behavior, resource requests/limits, HPA/VPA, node pool design, cost‑aware autoscaling (Cast AI, Karpenter, or equivalent).

Cloud and on‑premise infrastructure: GCP fluency, IaC (Terraform), CI/CD, and comfort operating in hy brid setups including on‑prem GPU clusters.

GPU workload understanding: familiarity with throughput profiling, batching, KV‑cache behaviour, inference server tuning, and GPU utilisation metrics.

Observability and reliability: metrics, traces, logs, SLOs, and the discipline to instrument systems properly rather than reactively.

FinOps mindset: demonstrated history of converting infrastructure choices into measurable cost outcomes.

Security baseline: able to take on platform‑security workstreams without requiring constant handoff to the DevOps team.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (Noida, BLR, India)
Senior Site Reliability Engineer (Noida, BLR, India)

Level AI • Dadri

Hybrid
INR 3,500,000 - 7,000,000
Senior Site Reliability Engineer (Noida, BLR, India)
Senior Site Reliability Engineer (Noida, BLR, India)

JobCubby • Bengaluru

On-site
INR 3,000,000 - 5,500,000
Senior Full Stack Engineer (Noida, BLR, India)
Senior Full Stack Engineer (Noida, BLR, India)

Level AI • Dadri, Bengaluru

On-site
INR 1,500,000 - 2,100,000
Principal Site Reliability Engineer
Principal Site Reliability Engineer

Namely • India

On-site
INR 1,500,000 - 2,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

SourcingXPress • Hyderabad

On-site
INR 3,000,000 - 5,000,000
Senior Full Stack Engineer (Noida, BLR, India)
Senior Full Stack Engineer (Noida, BLR, India)

Level AI • India

On-site
INR 1,200,000 - 2,400,000
Senior FullStack Engineer (Noida, BLR, India)
Senior FullStack Engineer (Noida, BLR, India)

StartX • Dadri, Bengaluru

Hybrid
INR 800,000 - 1,200,000
AI SRE/ AI Site Reliability Engineer
AI SRE/ AI Site Reliability Engineer

Tata Consultancy Services • Bengaluru

On-site
INR 1,800,000 - 2,800,000
Principal Software Engineer — Backend & Infrastructure - Noida / Bengaluru
Principal Software Engineer — Backend & Infrastructure - Noida / Bengaluru

Level AI • Dadri, Bengaluru

On-site
INR 6,000,000 - 9,000,000
Health coverage
Home-office allowance
Learning budget
+1
Senior Manager - Site Reliability Engineer|NR-2026-0246
Senior Manager - Site Reliability Engineer|NR-2026-0246

Media.net • Bengaluru

On-site
INR 6,000,000 - 8,000,000