Senior Site Reliability Engineer

Jobgether

India

Hybrid

INR 1,500,000 - 2,100,000

Full time

5 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Remote work in India
Flexible workplace
Growth opportunities
Collaborative environment

Job summary

Jobgether on behalf of a partner company seeks a Senior Site Reliability Engineer based in India. You will design and operate reliable, scalable infrastructure for distributed compute environments, focusing on automation, observability, and systems engineering.

You will collaborate across engineering, operations and support teams to improve service reliability, implement IaC, and modernize platforms across cloud and on-prem environments.

Qualifications

  • Bachelor’s degree in Computer Science, Engineering, or related technical discipline.
  • 6+ years of experience in Site Reliability Engineering, infrastructure or DevOps.
  • Strong Linux sysadmin expertise with production troubleshooting experience.
  • Solid networking knowledge including TCP/IP, DNS, routing, switching.
  • Hands-on with containers and Kubernetes for production systems.
  • Strong CI/CD experience using Jenkins and Git; observability with Prometheus/Grafana.
  • IaC and configuration management with Terraform/Ansible or similar.
  • Experience with cloud storage and distributed infrastructures.
  • Automation and scripting in Python, Bash, Go, or similar.
  • Excellent collaboration and communication skills across teams.

Responsibilities

  • Solve complex infrastructure and reliability challenges via troubleshooting and automation.
  • Deploy, operate, and maintain observability platforms and internal tools.
  • Collaborate with development, operations, and support teams to ensure reliability.
  • Provide technical guidance to engineers to meet reliability goals.
  • Investigate production issues across distributed systems and networks.
  • Modernize tooling and develop new applications for compute services.
  • Improve operations through IaC, monitoring, and repeatable practices.
  • Identify reliability risks and drive service-level improvements.

Skills

Linux administration
Networking fundamentals
Kubernetes
CI/CD & DevOps
Infrastructure as Code
Python / Bash / Go
Automation & scripting
Observability & monitoring
Collaboration & communication
Troubleshooting complex systems

Education

Bachelor’s degree in CS/Engineering or related

Tools

Jenkins
Git
Prometheus
Grafana
Terraform
Ansible / SaltStack

Job description

This position is listed on behalf of a partner company, who manages all applications and next steps. Our partner is looking for a Senior Site Reliability Engineer based in India.
This is a senior engineering role focused on building and operating reliable, scalable infrastructure and software for distributed compute environments. You will help solve complex reliability challenges through proactive troubleshooting, automation, observability, and systems engineering. The role combines infrastructure expertise with software development and close collaboration across operations, engineering, and support teams. You will design and maintain internal tooling and observability platforms that improve system visibility, performance, and service reliability. Working across a diverse technology landscape, you will contribute to modernizing existing systems while supporting the rollout of new applications and services. This opportunity is ideal for an experienced engineer who enjoys solving large-scale technical problems and improving the resilience of critical platforms.

Accountabilities
  • Solve complex infrastructure and reliability challenges through proactive troubleshooting, automation, systems programming, and structured root-cause analysis.
  • Deploy, operate, and maintain observability platforms and internal engineering tools that improve system visibility, reliability, and operational performance.
  • Partner with application development, operations, and support teams to ensure services are reliable, scalable, performant, and usable.
  • Provide technical guidance to engineers and developers, helping teams establish confidence that their services meet reliability and performance expectations.
  • Investigate and troubleshoot complex production issues across distributed systems, infrastructure, networking, and application environments.
  • Contribute to the modernization of existing tooling and the development of new applications supporting compute and distributed infrastructure services.
  • Improve operational processes through automation, infrastructure as code, monitoring, and repeatable engineering practices.
  • Collaborate across teams to identify reliability risks, strengthen service-level performance, and continuously improve operational excellence.
Requirements
  • Bachelor’s degree in Computer Science, Engineering, or a related technical discipline.
  • 6+ years of experience in Site Reliability Engineering, infrastructure engineering, DevOps, systems engineering, or a closely related field.
  • Strong Linux system administration expertise and practical experience troubleshooting production environments.
  • Solid understanding of networking fundamentals, including TCP/IP, DNS, routing and switching, and storage concepts.
  • Hands-on experience with containerized environments and Kubernetes, including operating, monitoring, and troubleshooting production systems.
  • Strong knowledge of CI/CD and DevOps practices, with practical experience using tools such as Jenkins, Git, Prometheus, and Grafana.
  • Experience implementing Infrastructure as Code and configuration management using Terraform, Ansible, SaltStack, or comparable technologies.
  • Familiarity with cloud storage systems and distributed infrastructure environments.
  • Strong automation and scripting capabilities using Python, Bash, Go, Rust, or similar programming languages.
  • Strong analytical and problem-solving skills, with the ability to investigate complex technical issues and develop reliable long-term solutions.
  • Excellent collaboration and communication skills, with the ability to work effectively across engineering, operations, and support teams.
  • Comfortable working in a fast-paced environment where systems and technologies continuously evolve.
Benefits
  • Remote work opportunity in India, with flexibility to work from home, an office, or a combination depending on role requirements.
  • Opportunity to work on large-scale distributed systems and complex Site Reliability Engineering challenges.
  • Exposure to modern technologies across cloud, compute, Kubernetes, observability, automation, and infrastructure engineering.
  • Collaboration with highly skilled engineering, operations, and application development teams.
  • Opportunity to influence system reliability, scalability, monitoring, and operational excellence across critical services.
  • Support for employee health, well-being, financial security, and life beyond work through comprehensive benefits.
  • Professional growth through exposure to diverse technologies, modernization initiatives, and large-scale engineering environments.
  • Flexible workplace approach designed to support productive and effective ways of working.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer 3 (Remote - India)
Site Reliability Engineer 3 (Remote - India)

Jobgether • India

Remote
INR 1,800,000 - 2,500,000
Competitive salary and bonuses
Flexible remote work
Paid time off and wellbeing days
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Falabella India • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hdfc Bank • Bengaluru

On-site
INR 2,500,000 - 4,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

BuildxPartners • Bengaluru Urban

Hybrid
INR 2,400,000 - 4,200,000
SRE Engineer @ Investment Banking | Mumbai
SRE Engineer @ Investment Banking | Mumbai

Net Connect Global • Bengaluru, Mumbai

Hybrid
INR 1,800,000 - 2,400,000
Site Reliability Engineer
Site Reliability Engineer

Arch Systems • Hyderabad

On-site
INR 2,800,000 - 4,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Infosys • Hyderabad

On-site
INR 1,400,000 - 2,200,000
Site Reliability Engineer II
Site Reliability Engineer II

Jobgether • India

Remote
INR 1,500,000 - 2,100,000
Remote work in India
Flexible working arrangements
Health & well-being benefits
+3
Site Reliability Engineer
Site Reliability Engineer

Jobtailor • Thiruvananthapuram

On-site
INR 2,800,000 - 5,200,000
Hybrid work environment
Life Insurance
Paid holidays
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Pune District

On-site
INR 1,200,000 - 1,800,000
Health insurance
Flexible working hours
Training opportunities