Site Reliability Engineer

TVS Next

Chennai District

Hybrid

INR 1,200,000 - 1,600,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work model
Comprehensive family health insurance
Opportunities for continuous learning
Clear career path and manager guidance
Mentorship within diverse groups

Job summary

TVS Next in Chennai is looking for a dedicated Site Reliability Engineer to help design and maintain high-availability cloud infrastructure. Your role will involve managing AWS services, optimizing databases, and implementing CI/CD pipelines.

We seek candidates with at least 5 years of experience in cloud environments, hands-on AWS expertise, and a strong understanding of Infrastructure as Code. Join us for a hybrid work model that promotes work-life balance and offers comprehensive benefits.

Qualifications

  • 5+ years of experience designing and managing scalable cloud infrastructure.
  • Strong hands-on expertise in AWS cloud services and cloud-native architecture.
  • Experience with CI/CD automation pipelines using modern DevOps practices.

Responsibilities

  • Design and manage scalable AWS cloud infrastructure for business applications.
  • Optimize database infrastructure and develop Infrastructure as Code frameworks.
  • Implement CI/CD pipelines for deployment efficiency and reliability.

Skills

AWS cloud services
Containerized environments
Relational database management
Infrastructure as Code (IaC)
CI/CD automation
Networking and troubleshooting
Site Reliability Engineering principles

Education

Bachelor’s or Master’s degree in Computer Science, Engineering, Information Technology

Tools

Terraform
AWS CloudFormation
AWS ECS
AWS Fargate
PostgreSQL

Job description

We are looking for Site Reliability Engineer - Chennai

About The Job

Join our high‑performance Cloud Engineering team and play a critical role in designing, automating, and maintaining highly available, scalable, and resilient cloud infrastructure platforms that power enterprise‑grade applications and digital experiences.

What You’ll Do
  • Design, implement, and manage scalable AWS cloud infrastructure to support mission‑critical business applications.
  • Build, deploy, and maintain containerized workloads using AWS ECS and AWS Fargate.
  • Manage and optimize relational database infrastructure leveraging AWS RDS and PostgreSQL.
  • Develop and maintain Infrastructure as Code (IaC) frameworks to enable automated, repeatable, and consistent infrastructure deployments.
  • Design, implement, and optimize CI/CD pipelines to improve deployment efficiency, reliability, and release velocity.
  • Implement and optimize high‑performance caching strategies to improve application responsiveness, scalability, and overall system performance.
  • Ensure infrastructure availability, reliability, scalability, and security across production and non‑production environments.
  • Monitor platform health, system performance, and application availability while proactively identifying and addressing potential issues.
  • Troubleshoot and resolve complex infrastructure, networking, and platform‑related issues across distributed systems.
  • Perform root cause analysis for production incidents and drive preventive actions to improve system reliability.
  • Collaborate closely with software engineering, architecture, DevOps, and QA teams to support cloud‑native application delivery.
  • Establish and implement reliability engineering best practices including observability, automation, incident management, and capacity planning.
  • Document operational procedures, architecture decisions, and best practices to ensure knowledge sharing and operational excellence.
  • Contribute to continuous improvement initiatives focused on system reliability, operational efficiency, and infrastructure modernization.
What We Seek In You
  • 5+ years of experience designing, implementing, managing, and troubleshooting scalable cloud infrastructure environments.
  • Strong hands‑on expertise in AWS cloud services and cloud‑native architecture patterns.
  • Proven experience working with containerized environments using AWS ECS and AWS Fargate.
  • Strong experience managing relational database platforms including AWS RDS and PostgreSQL.
  • Advanced proficiency in Infrastructure as Code (IaC) tools such as Terraform, AWS CloudFormation, or equivalent technologies.
  • Extensive experience designing and implementing CI/CD automation pipelines using modern DevOps practices and tools.
  • Strong expertise implementing and optimizing high‑performance caching layers and distributed caching solutions.
  • Solid understanding of cloud architecture principles including high availability, scalability, fault tolerance, and disaster recovery.
  • Strong troubleshooting and analytical skills with the ability to diagnose issues across infrastructure, networking, and applications.
  • Experience working in Linux‑based environments and cloud‑native ecosystems.
  • Strong understanding of Site Reliability Engineering principles including observability, monitoring, automation, and incident management.
  • Experience with monitoring, logging, and observability tools for proactive system management.
  • Excellent communication, stakeholder management, and collaboration skills.
  • Experience working within Agile and DevOps delivery environments.
Preferred Qualifications
  • Hands‑on experience deploying and managing workloads within AWS China regions.
  • Strong understanding of network‑level troubleshooting including VPC, DNS, routing, load balancers, security groups, and connectivity diagnostics.
  • Experience troubleshooting distributed systems and high‑scale cloud‑native applications.
  • Familiarity with highly scalable API platforms and backend services supporting enterprise applications.
  • Exposure to CDN, caching architectures, and edge delivery strategies is preferred.
  • Knowledge of Kubernetes, Docker, and modern container orchestration technologies is an added advantage.
  • Experience with performance testing, load testing, and capacity planning methodologies is preferred.
  • Bachelor’s or Master’s degree in Computer Science, Engineering, Information Technology, or a related discipline.
Benefits
  • Hybrid work model promoting work‑life balance.
  • Access to comprehensive family health insurance coverage.
  • Opportunities for continuous learning and upskilling through internal training programs.
  • Clear career path and guidance from managers via ongoing feedback sessions.
  • Engagement opportunities with customers, product managers, and leadership.
  • Supportive community and mentorship within diverse groups of interest.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Arch Systems • Hyderabad

On-site
INR 2,800,000 - 4,200,000
Site Reliability Engineer 3 (Remote - India)
Site Reliability Engineer 3 (Remote - India)

Jobgether • India

Remote
INR 1,800,000 - 2,500,000
Competitive salary and bonuses
Flexible remote work
Paid time off and wellbeing days
+3
Senior Site Reliability Engineer (SRE) – AWS
Senior Site Reliability Engineer (SRE) – AWS

Sailssoftware • Visakhapatnam

On-site
INR 1,500,000 - 2,500,000
Site Reliability Engineer
Site Reliability Engineer

MNR Solutions Pvt. Ltd. • Bengaluru

On-site
INR 900,000 - 1,500,000
Software Engineer (Site Reliability Engineer)
Software Engineer (Site Reliability Engineer)

H&R Block India • Thiruvananthapuram

On-site
INR 600,000 - 800,000
Site Reliability Engineer
Site Reliability Engineer

Plume • Hyderabad

On-site
INR 2,500,000 - 5,200,000
Site Reliability Engineer
Site Reliability Engineer

SourcingXPress • Mumbai

On-site
INR 800,000 - 1,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

F-Prime Capital • Pune District

On-site
INR 1,500,000 - 2,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Fairygodboss • Chennai District

On-site
INR 1,000,000 - 2,000,000
Staff Site Reliability Engineer
Staff Site Reliability Engineer

United States Digital Space LLC • Bengaluru

Hybrid
INR 6,000,000 - 12,000,000