Senior Lead Site Reliability Engineer

FIS

Pune District

On-site

INR 3,500,000 - 6,000,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

FIS seeks a Senior Lead Site Reliability Engineer to define, build, and operate always-on, low-latency, and secure payment platforms powering large-scale financial transactions.

The role focuses on automating, designing, and supporting highly available, mission-critical systems across distributed architectures, cloud, and platform automation. You will lead complex initiatives from design to production and collaborate across engineering teams.

Qualifications

  • 10 to 15 years of IT experience.
  • Deep software engineering expertise with distributed systems.
  • Strong CI/CD experience with major tools.
  • Expertise in observability and reliability engineering.
  • Strong cloud platform experience (AWS/Azure/GCP).
  • Hands-on Linux/Windows, and databases in enterprise stacks.

Responsibilities

  • Design reliability architecture and standards across services and infrastructure.
  • Build and evolve observability platforms (metrics/logs/traces/SLOs/SLIs).
  • Improve SRE practices like error budgets and resilience testing.
  • Develop automation and self-service platforms to reduce toil.
  • Design, implement, and optimize CI/CD and release strategies.
  • Troubleshoot production issues and perform root-cause analysis.
  • Collaborate with teams to improve reliability and performance.

Skills

CI/CD pipelines
Observability tooling
Cloud platforms
Linux systems
Troubleshooting
Distributed systems

Tools

Azure DevOps
GitHub Actions
Jenkins
Harness

Job description

Senior Lead Site Reliability Engineer
About The Role
  • We are hiring a Senior Lead Site Reliability Engineer to define, build, and operate always-on, low-latency, and highly secure payment platforms that power large-scale financial transactions.
  • This is a senior hands-on engineering role focused on designing, automating, and supporting highly available, mission-critical platforms. You will work at the intersection of distributed systems, cloud platforms, observability, and reliability engineering, helping to improve platform stability, performance, scalability, and operational efficiency.
  • You will collaborate with engineering, development, and infrastructure teams to implement reliable solutions, automate operational processes, and resolve complex production issues while remaining deeply involved in the day-to-day technical work.
  • The ideal candidate is a seasoned engineer with strong problem-solving skills, a passion for automation, and the ability to independently drive complex technical initiatives from design through production.
What you be doing
  • Design and implement reliability architecture and engineering standards across services, platforms, and infrastructure, ensuring solutions are scalable, resilient, and operationally efficient.
  • Design, build, and evolve enterprise-grade observability platforms (metrics, logs, traces, SLOs/SLIs) that provide actionable insights into system health, customer experience, and business impact.
  • Implement and continuously improve SRE practices including error budgets, capacity modeling, resilience testing, graceful degradation, and operational readiness.
  • Architect, build, and enhance automation and self-service platforms that eliminate toil, reduce operational risk, and enable safe, frequent production releases.
  • Design, implement, and optimize CI/CD and release engineering solutions, including secure pipelines, deployment automation, quality gates, and rollback strategies.
  • Troubleshoot and resolve complex production issues, performing deep root-cause analysis and implementing long-term reliability improvements.
  • Collaborate with engineering, platform, and infrastructure teams to improve system reliability, scalability, performance, and operational excellence.
What You Bring
  • 10 to 15 Years of IT Experience.
  • Deep software engineering expertise with a proven track record of designing, building, and operating large-scale, distributed, API-driven systems in production.
  • Strong CI/CD experience with tools such as Azure DevOps, GitHub Actions, Jenkins, Harness, or equivalent, including pipeline design, deployment automation, and release engineering best practices.
  • Expertise in observability, alerting, and reliability engineering, using tools such as Prometheus, Grafana, Datadog, Splunk, ELK, or equivalent ecosystems.
  • Strong command of cloud platforms and open systems (AWS, Azure, or GCP), including infrastructure-as-code, platform automation, and cloud-native design patterns.
  • Hands-on experience across Linux (RHEL), Windows, databases (e.g. SQL Server, Oracle RDBMS), and complex enterprise technology stacks with strong troubleshooting and debugging skills.
  • Strong experience designing and operating high-availability, mission-critical platforms with a focus on reliability, performance, scalability, and automation.
Added bonus if you have:
  • Strong automation and scripting skills using Python, Bash, Ansible, PowerShell, or similar tools, with additional experience in C#/.NET considered a strong plus.
  • Prior ownership of reliability strategy or platform initiatives spanning multiple teams or business units.
  • Experience modernizing legacy financial systems into cloud-native or hybrid architectures with a focus on resilience and compliance.
What we offer you
  • A multifaceted job with a high degree of responsibility and a broad spectrum of opportunities
  • A modern, international work environment and a dedicated and motivated team
  • A broad range of professional education and personal development possibilities FIS is your final career step!
  • A competitive salary and benefits
  • A variety of career development tools, resources and opportunities
Privacy Statement

FIS is committed to protecting the privacy and security of all personal information that we process in order to provide services to our clients. For specific information on how FIS protects personal information online, please see the .

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Site Reliability Engineer
Lead Site Reliability Engineer

FIS • Bengaluru

On-site
INR 350,000 - 700,000
Support Engineer Enterprise Platforms
Support Engineer Enterprise Platforms

FIS • Bengaluru

Hybrid
INR 3,500,000 - 5,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

BayOne Solutions • Hyderabad

On-site
INR 3,000,000 - 6,000,000
Senior DevOps Engineer(6+ years' experience in DevOps, AWS, Kubernetes and Docker)
Senior DevOps Engineer(6+ years' experience in DevOps, AWS, Kubernetes and Docker)

FIS • Pune District

On-site
INR 2,800,000 - 4,400,000
Competitive salary
Inclusive, diverse environment
Learning and development opportunities
Site Reliability Engineer (Java,Unix,Dynatrace and Splunk)
Site Reliability Engineer (Java,Unix,Dynatrace and Splunk)

FIS • Pune District

On-site
INR 1,000,000 - 1,400,000
Opportunity in a leading FinTech company
Professional development possibilities
High degree of responsibility
Lead Site Reliability Engineer SRE
Lead Site Reliability Engineer SRE

FIS • Chennai District

Hybrid
INR 3,500,000 - 5,500,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Impronics Technologies • Gurugram District

On-site
INR 1,500,000 - 2,500,000
Senior Enterprise Platform Support Engineer AI & Cloud
Senior Enterprise Platform Support Engineer AI & Cloud

FIS • Bengaluru

On-site
INR 3,000,000 - 4,500,000
Senior Full Stack UI Engineer
Senior Full Stack UI Engineer

FIS • Pune District

On-site
INR 1,800,000 - 3,000,000
Impactful fintech work
Learning and growth opportunities
Inclusive, diverse environment
+2
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

FIS • India

Hybrid
INR 3,000,000 - 5,000,000
Hybrid work schedule
Competitive salary
Growth opportunities