Site Reliability Engineer

LambdaTest, Inc.

Dadri

On-site

INR 900,000 - 1,300,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

LambdaTest, Inc. is seeking a Platform Engineer to own and scale our cloud infrastructure. You will work on production-grade systems, writing automation, and driving reliability across live services.

Expect hands-on experience with AWS, Kubernetes, and containerized workloads. You will build robust CI/CD pipelines with Jenkins, ArgoCD, and Helm, while expanding observability with Prometheus and Grafana.

Qualifications

  • Proven hands-on DevOps or Cloud Infrastructure experience.
  • Production AWS experience (EKS, SQS, ECR, Route 53, ALB/NLB).
  • Production Docker and Kubernetes experience.
  • Hands-on with at least one APM tool (New Relic, Sumo Logic, Prometheus, Grafana).
  • Real production incident experience with root-cause articulation.

Responsibilities

  • Own incidents end-to-end and build custom automation on live production systems.
  • Own and scale services on AWS — EKS, EC2, SQS, ECR, Route 53, ALB/NLB.
  • Manage production workloads on EKS; implement Karpenter and KEDA for autoscaling.
  • Build and optimize CI/CD pipelines using Jenkins, ArgoCD, Argo, and Helm.
  • Observe and debug production using Prometheus, Grafana, New Relic, or Sumo Logic.

Skills

DevOps
Cloud infrastructure
Kubernetes
Containers
Observability
Incident management
Programming
System thinking

Tools

AWS
Docker
Jenkins
Terraform
ArgoCD
Helm
Istio
Prometheus
Grafana

Job description

LambdaTest is a high-growth SaaS platform powering millions of test executions globally with 100% YoY traffic growth. Join our Platform Engineering team to own and scale the cloud infrastructure that keeps our systems fast, reliable, and production-ready.

The Role: What You'll Do

  • This is a genuine DevOps/SRE role — not config management or ticket ops. You'll write code, own incidents end-to-end, and build custom automation solutions on live production systems.
  • Cloud Operations: Own and scale services on AWS — EKS, EC2, SQS, ECR, Route 53, ALB/NLB.
  • Kubernetes & Containers: Manage production workloads on EKS; implement Karpenter and KEDA for event-driven and node autoscaling.
  • CI/CD: Build and optimize pipelines using Jenkins, Docker Buildx, ECR caching, ArgoCD, and Helm.
  • Observability & Incident Management: Own APM, monitoring, and production debugging using New Relic, Sumo Logic, Prometheus, or Grafana.
  • Automation & Scripting: Write custom code for logical and automation solutions — not just run commands on machines.
  • IaC: Use Terraform for provisioning; understand state management and concurrent apply risks in team environments.
You’ll Thrive Here If You Have | Must-Haves:
  • Experience: 1–3 years of hands-on DevOps or Cloud Infrastructure experience.
  • Cloud: Production experience on AWS — EKS, SQS, ECR, Route 53, ALB/NLB.
  • Containers: Docker and Kubernetes in production environments.
  • Observability: Hands-on with at least one APM tool — New Relic, Sumo Logic, Prometheus, or Grafana.
  • Incident Management: Real production incident experience — must articulate root cause, not just symptoms.
  • Coding: Genuine programming ability in any language — this role builds custom solutions, not just runs CLI commands.
  • Mindset: Bridges dev and DevOps thinking; can explain why architectural choices were made, not just what was done.
Good to Have:
  • Golang or Java — backend coding ability is a strong plus.
  • KEDA and Karpenter — event-driven and node autoscaling experience.
  • ArgoCD, Helm, or Istio — GitOps and service mesh exposure.
  • Kafka or SQS — messaging and queue configuration knowledge.
  • System design awareness at the services layer — how services interact, not just surface-level ops.
What We Offer

Direct exposure to production systems at scale, ownership from day one, and a clear growth path in Platform and SRE Engineering.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Innodata Inc. • India

On-site
INR 2,400,000 - 4,000,000
Lead DevOps Engineer
Lead DevOps Engineer

Lenskart • Gurugram District

On-site
INR 1,200,000 - 2,400,000
Site Reliability Engineer
Site Reliability Engineer

DeepIQ • Hyderabad

On-site
INR 1,200,000 - 2,100,000
DevOps Engineer
DevOps Engineer

NAVVYASA CONSULTING PRIVATE LIMITED • Gurugram District

On-site
INR 800,000 - 1,200,000
Devops Engineer
Devops Engineer

SAIGroup • Bengaluru

On-site
INR 1,200,000 - 2,000,000
Competitive compensation
Equity
Benefits
Site Reliability Engineer
Site Reliability Engineer

Insight Global • Bengaluru

On-site
INR 2,500,000 - 5,000,000
Lead SDE - DevOps
Lead SDE - DevOps

Flourish Ventures • Chennai District

On-site
INR 2,000,000 - 3,000,000
Inclusive and people-first culture
Health & wellness programs
Comprehensive medical insurance
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Headout • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Sr DevOps Engineer
Sr DevOps Engineer

Clarity RCM • Chennai District

On-site
INR 3,200,000 - 5,200,000
Cloud Operations Lead – SRE / DevOps / Platform Engineering
Cloud Operations Lead – SRE / DevOps / Platform Engineering

PeoplePilot • Pune District

On-site
INR 2,600,000 - 5,200,000