Site Reliability Engineer - Career

Equifax

Pune District

On-site

INR 2,400,000 - 4,200,000

Full time

9 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Equifax in Pune is seeking an experienced SRE/DevOps engineer to manage uptime across cloud-native and hybrid architectures. You will implement infrastructure as code with Terraform, build CI/CD pipelines with Jenkins, and automate deployment of service changes to production.

Responsibilities include incident triage, runbook improvement, blameless postmortems, and mentoring other engineers while ensuring security and reliability in a regulated enterprise environment.

Qualifications

  • BS degree in Computer Science or related technical field involving coding, or equivalent job experience.
  • 5-7 years of experience in software engineering, systems administration, database administration, and networking.
  • 2+ years of experience developing and/or administering software in public cloud.
  • Proficiency with continuous integration and continuous delivery tooling and practices.
  • System administration skills, including automation and orchestration of Linux/Windows using Terraform, Chef, Ansible and/or containers (Docker, Kubernetes, etc.).
  • Experience in languages such as Python, Bash, Java, Go JavaScript and/or node.js.
  • Experience in monitoring infrastructure and application uptime and availability to ensure functional and performance objectives.

Responsibilities

  • Manage system(s) uptime across cloud-native (AWS, GCP) and hybrid architectures.
  • Build infrastructure as code (IAC) patterns using Terraform and cloud SDKs.
  • Build CI/CD pipelines for build, test and deployment using Jenkins and cloud-native toolchains.
  • Build automated tooling to deploy service requests and runbooks to manage production changes.
  • Triage incidents, improve run books, and lead blameless postmortems to reduce MTTR.
  • Lead availability improvements and drive remediation actions.

Job description

What You'll Do
  • Manage system(s) uptime across cloud-native (AWS, GCP) and hybrid architectures.
  • Build infrastructure as code (IAC) patterns that meet security and engineering standards using one or more technologies (Terraform, scripting with cloud CLI, and programming with cloud SDK).
  • Build CI/CD pipelines for build, test and deployment of application and cloud architecture patterns, using platform (Jenkins) and cloud-native toolchains.
  • Build automated tooling to deploy service requests to push a change into production. Build runbooks that are comprehensive and detailed to manage detect, remediate and restore services.
  • Solve problems and triage complex distributed architecture service maps. On call for high severity application incidents and improving run books to improve MTTR
  • Lead availability blameless postmortem and own the call to action to remediate recurrences.
What Experience You Need
  • BS degree in Computer Science or related technical field involving coding (e.g., physics or mathematics), or equivalent job experience required
  • 5-7 years of experience in software engineering, systems administration, database administration, and networking
  • 2+ years of experience developing and/or administering software in public cloud
  • Cloud Certification Strongly Preferred
  • Proficiency with continuous integration and continuous delivery tooling and practices
  • System administration skills, including automation and orchestration of Linux/Windows using Terraform, Chef, Ansible and/or containers (Docker, Kubernetes, etc.)
  • Demonstrable cross-functional knowledge with systems, storage, networking, security and databases
  • Experience in languages such as Python, Bash, Java, Go JavaScript and/or node.js
  • Experience in monitoring infrastructure and application uptime and availability to ensure functional and performance objectives
What Could Set You Apart
  • You have expertise designing, analyzing and troubleshooting large-scale distributed systems.
  • You take a system problem-solving approach, coupled with strong communication skills and a sense of ownership and drive
  • Kubernetes (CKA, CKAD) or cloud certifications.
  • You are passionate for automation with a desire to eliminate toil whenever possible
  • You’ve built software or maintained systems in a highly secure, regulated or compliant industry
  • You thrive in and have experience and passion for working within a DevOps culture and as part of a team
  • BS in Computer Science or related field.
  • 2+ years of experience developing and/or administering software in public cloud
  • 5+ years programming experience (Python, Bash/Shell Script, Java, Go, etc.).
  • 3+ years of experience monitoring infrastructure and application performance.
  • 5+ years experience of system administration skills, including automation and orchestration of Linux/Windows using Terraform, Chef, Ansible and/or containers (Docker, Kubernetes, etc.)
  • 5+ years experience working with continuous integration and continuous delivery tooling and practices
  • Kubernetes: Design, deploy, and manage production-ready Kubernetes clusters.
  • Cloud Infrastructure: Build and maintain scalable infrastructure on GCP using tools like Terraform.
  • Performance: Identify and resolve performance bottlenecks in applications and infrastructure.
  • Observability: Implement monitoring and logging to proactively detect and resolve issues.
  • Incident Response: Participate in on-call rotations, troubleshooting and resolving production incidents.
  • Collaboration: Promote reliability best practices and ensure smooth deployments.
  • Automation: Build CI/CD pipelines, automated tooling, and runbooks.
  • Problem Solving: Triage complex issues, lead blameless postmortems, and drive remediation.
  • Mentorship: Guide and mentor other SREs.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Smart Ims • Bengaluru

Hybrid
INR 1,200,000 - 2,000,000
Senior Associate Site Reliability Engineer
Senior Associate Site Reliability Engineer

NTT DATA BUSINESS SOLUTIONS • Hyderabad

On-site
INR 1,400,000 - 2,000,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Site Reliability Engineering (SRE)
Site Reliability Engineering (SRE)

Lyzr AI • Bengaluru

Hybrid
INR 1,000,000 - 2,000,000
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

Skillventory • Kamrup Metropolitan

On-site
INR 3,500,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Infosys • Bengaluru

On-site
INR 900,000 - 1,500,000
Senior SRE Engineer
Senior SRE Engineer

Epam Systems • Chennai District

On-site
INR 2,500,000 - 4,000,000
Lead SRE
Lead SRE

Cvent, Inc. • Gurugram District

On-site
INR 4,000,000 - 8,000,000
SRE - AWS, GCP & Azure
SRE - AWS, GCP & Azure

PibyThree • Navi Mumbai

On-site
INR 1,200,000 - 1,800,000
SRE - AWS, GCP & Azure
SRE - AWS, GCP & Azure

PibyThree • Thane

On-site
INR 1,200,000 - 1,500,000