Lead Software Engineer - Site Reliability

jobr.pro

Chennai District

On-site

INR 3,000,000 - 5,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Freshworks is seeking a Lead Site Reliability Engineer (SRE) in India to design resilient systems, automate recovery, and maintain observability at scale. You will partner with engineering, platform, and product teams to raise reliability and performance while shifting left on quality.

You will lead incident response and postmortems, build automated pipelines, and champion best practices across services to ensure 99.99%+ uptime and scalable infrastructure.

Qualifications

  • 7–12 years in SRE, DevOps, or Production Engineering roles.
  • Coding proficiency to develop clear, efficient code.
  • Linux expertise for system administration and troubleshooting.
  • Containerization & orchestration with Docker and Kubernetes for deployment.
  • CI/CD management across design, implementation, and maintenance.
  • Security and compliance in infrastructure.
  • High availability and scalable distributed systems.
  • Infrastructure as Code (IaC) tooling and automation.
  • Disaster Recovery and High Availability knowledge and practice.
  • Observability: monitoring, logging, and tracing.

Responsibilities

  • Design and implement tools to improve availability, latency, scalability, and system health.
  • Define SLIs/SLOs, manage error budgets, and drive performance engineering efforts.
  • Build and maintain automated monitoring, alerting, and remediation pipelines.
  • Collaborate with engineering teams to improve reliability by design.
  • Lead incident response, root cause analysis, and blameless postmortems.
  • Champion observability across services—logs, metrics, traces.
  • Contribute to infrastructure architecture, automation, and reliability roadmaps.
  • Advocate for SRE best practices across teams and functions.

Skills

SRE
Linux
Docker
Kubernetes
CI/CD
IaC
Observability
System design
Security
Disaster recovery
Automation

Education

Bachelor’s degree in CS/Engineering

Tools

Terraform
Ansible

Job description

Job Description

At Freshworks, uptime is sacred. As a Lead Site Reliability Engineer (SRE), you'll be the engineer behind the curtain—designing for resilience, automating recovery, and ensuring our systems stay fast, stable, and observable at scale. You’ll partner closely with engineering, platform, and product teams to shift reliability left and set the standard for performance and availability.

If you live for clean telemetry, root cause resolution, and engineering chaos into confidence, this is your playground.

Responsibilities
  • Design and implement tools to improve availability, latency, scalability, and system health.
  • Define SLIs/SLOs, manage error budgets, and drive performance engineering efforts.
  • Build and maintain automated monitoring, alerting, and remediation pipelines.
  • Collaborate with engineering teams to improve reliability by design.
  • Lead incident response, root cause analysis, and blameless postmortems.
  • Champion observability across services—logs, metrics, traces.
  • Contribute to infrastructure architecture, automation, and reliability roadmaps.
  • Advocate for SRE best practices across teams and functions.
Qualifications
  • 7–12 years of experience in SRE, DevOps, or Production Engineering roles.
  • Coding Proficiency: Develop clear, efficient, and well-structured code.
  • Linux Expertise: In-depth knowledge of Linux for system administration and advanced troubleshooting.
  • Containerization & Orchestration: Practical experience with Docker and Kubernetes for application deployment and management.
  • CI/CD Management: Design, implement, and maintain Continuous Integration and Continuous Delivery pipelines.
  • Security & Compliance: Understand security best practices and compliance in infrastructure.
  • High Availability & Scalability: Design and implement highly available, scalable, and resilient distributed systems.
  • Infrastructure as Code (IaC) & Automation: Proficient in IaC tools and automating infrastructure provisioning and management.
  • Disaster Recovery (DR) & High Availability (HA): Deep knowledge and practical experience with various DR and HA strategies.
  • Observability: Implement and utilize monitoring, logging, and tracing tools for system health.
  • System Design (Distributed Systems): Design complex distributed systems with a focus on reliability and operations.
  • Problem-Solving & Troubleshooting: Excellent analytical and diagnostic skills for resolving complex system issues.
  • Degree in Computer Science, Engineering, or related field.
  • Experience building and scaling services in production with high uptime targets (99.99%+).
  • Clear track record of reducing incident frequency and improving response metrics (MTTD/MTTR).
  • Strong communicator who thrives in high-pressure environments.
  • Passionate about automation, chaos engineering, and making things just work.
Additional Information

At Freshworks, we have fostered an environment that enables everyone to find their true potential, purpose, and passion, welcoming colleagues of all backgrounds, genders, sexual orientations, religions, and ethnicities. We are committed to providing equal opportunity and believe that diversity in the workplace creates a more vibrant, richer environment that boosts the goals of our employees, communities, and business. Fresh vision. Real impact. Come build it with us.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Synechron • Bengaluru, Hyderabad

Hybrid
INR 4,200,000 - 6,300,000
Resilience and Reliability Engineer
Resilience and Reliability Engineer

EY • Pune District, Gurugram District, Bengaluru

Hybrid
INR 1,800,000 - 2,800,000
Lead SRE
Lead SRE

United States Digital Space LLC • Karnataka

On-site
INR 900,000 - 1,400,000
Site Reliability Engineer
Site Reliability Engineer

Spot Your Leaders & Consulting • Pune District

On-site
INR 2,500,000 - 4,000,000
Senior Software Engineer (Site Reliability Engineering)
Senior Software Engineer (Site Reliability Engineering)

SentiLink • Bengaluru

Hybrid
INR 1,200,000 - 1,800,000
Employer paid group health insurance
401(k) plan with employer match
Flexible paid time off
+2
Site Reliability Engineer
Site Reliability Engineer

SourcingXPress • Mumbai

On-site
INR 800,000 - 1,200,000
Site Reliability Engineer(SRE)
Site Reliability Engineer(SRE)

MetaForgeIT • Hyderabad

On-site
INR 1,000,000 - 1,500,000
Site Reliability Engineer II
Site Reliability Engineer II

United States Digital Space LLC • Karnataka

On-site
INR 800,000 - 1,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Five9 • Bengaluru

On-site
INR 4,000,000 - 7,000,000