Lead Site Reliability Engineer

fis

India

On-site

INR 4,000,000 - 6,500,000

Full time

7 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

FIS is seeking a Senior Lead Site Reliability Engineer to define, build, and operate highly secure payment platforms for large-scale financial transactions. This senior role sits at the intersection of distributed systems, cloud platforms, and reliability architecture, guiding multi-team efforts while remaining hands-on with complex reliability challenges.

You will own reliability outcomes, define standards, and lead incident response with a bias toward scalable, automated solutions, partnering

Qualifications

  • Deep software engineering experience building and operating large-scale distributed systems.
  • Proven expertise in observability and reliability engineering using modern tooling.
  • Strong cloud platform experience (AWS, Azure or GCP) and IaC practices.
  • Experience in regulated environments (Payments/FinTech) emphasizing availability and security.
  • Hands-on with Linux (RHEL), Windows and complex enterprise stacks and troubleshooting.

Responsibilities

  • Own and drive reliability outcomes at scale for real-time payment platforms with strict SLAs/SLOs.
  • Define reliability architecture and standards across services and platforms.
  • Design and evolve observability platforms (metrics, logs, traces, SLOs/SLIs) for health insights.
  • Lead incident management and drive root-cause analysis and long-term fixes.
  • Promote SRE practices including error budgets, capacity modeling, resilience testing, and readiness.
  • Architect automation and self-service platforms to enable frequent production releases.
  • Collaborate with leaders to influence cloud strategy, DR posture, and platform evolution.
  • Mentor senior engineers to raise reliability and maturity.

Skills

Software engineering
Observability
Cloud platforms
Linux
Windows
Incident management
CI/CD
Python

Tools

Prometheus
Grafana
Datadog
Splunk
ELK

Job description

About The Role

We are hiring a Senior Lead Site Reliability Engineer to define, build, and operate always-on, low-latency , and highly secure payment platforms that power large-scale financial transactions.

This is a senior technical role, not a pure operations position. You will operate at the intersection of distributed systems engineering, cloud platforms, and reliability architecture, setting technical direction and driving reliability outcomes across mission-critical, regulated systems in Payments and FinTech.

You will work across multiple teams and domains, influencing architecture, engineering practices, and operational maturity while remaining hands-on with the most complex reliability challenges.

What You Will Do
  • Own and drive reliability outcomes at scale for real-time, distributed payment and transaction processing platforms with strict SLAs, SLOs, and regulatory requirements.
  • Define reliability architecture and standards across services, platforms, and infrastructure-shaping how systems are designed, deployed, observed, and operated.
  • Design and evolve enterprise-grade observability platforms (metrics, logs, traces, SLOs/SLIs) that provide actionable insights into system health, customer experience, and business impact.
  • Lead and coordinate response to high-severity production incidents , acting as a technical authority during major events and driving deep root-cause analysis and long-term systemic fixes.
  • Set strategy and drive adoption of SRE best practices including error budgets, capacity modeling, resilience testing, graceful degradation, and operational readiness.
  • Architect automation and self-service platforms that eliminate toil, reduce operational risk, and enable safe, frequent production releases across teams.
  • Partner with senior engineering, product, and platform leaders to influence architectural decisions, cloud migration strategy, disaster recovery posture, and long-term platform evolution.
  • Mentor senior engineers and technical leads , raising the overall reliability and operational maturity of the organization.
What You Bring
  • Deep software engineering expertise with a proven track record of building and operating large-scale, distributed, API-driven systems in production.
  • Expertise in observability, alerting, and reliability engineering , using tools such as Prometheus, Grafana, Datadog, Splunk, ELK, or equivalent ecosystems.
  • Strong command of cloud platforms and open systems ( AWS, Azure, or GCP ), including infrastructure-as-code , platform automation, and cloud-native design patterns.
  • Significant experience running mission-critical systems in Payments, FinTech, Banking, or similarly regulated environments , where availability, correctness, and security are non-negotiable.
  • Hands-on experience across Linux (RHEL), Windows , databases (e.g., Oracle RDBMS), and complex enterprise stacks with strong system-level troubleshooting skills.
  • Demonstrated leadership in incident management, post-incident reviews, and continuous reliability improvement , with the ability to influence behavior and standards across teams.
  • Ability to operate effectively at Staff level scope-solving ambiguous problems, making trade-offs, and driving alignment across multiple teams and stakeholders.
Added Advantage
  • Strong automation and scripting skills using Python, Bash, Ansible, or similar tools.
  • Experience building or scaling CI/CD platforms and release automation in high-risk production environments.
  • Prior ownership of reliability strategy or platform initiatives spanning multiple teams or business units.
  • Experience modernizing legacy financial systems into cloud-native or hybrid architectures with a focus on resilience and compliance.
Why Join Us
  • Work on high-impact payment platforms operating at massive scale, where milliseconds and reliability directly affect real-world commerce.
  • Play a Staff-level role in defining reliability strategy for systems that cannot fail.
  • Join a culture that values engineering excellence, technical leadership, automation, and continuous learning.
Privacy Statement

FIS is committed to protecting the privacy and security of all personal information that we process in order to provide services to our clients. For specific information on how FIS protects personal information online, please see the Online Privacy Notice .

Sourcing Model

Recruitment at FIS works primarily on a direct sourcing model; a relatively small portion of our hiring is through recruitment agencies. FIS does not accept resumes from recruitment agencies which are not on the preferred supplier list and is not responsible for any related fees for resumes submitted to job postings, our employees, or any other part of our company.

#pridepass

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead Site Reliability Engineer
Lead Site Reliability Engineer

FIS • Bengaluru

On-site
INR 350,000 - 700,000
Senior Lead Site Reliability Engineer
Senior Lead Site Reliability Engineer

FIS • Pune District

On-site
INR 3,500,000 - 6,000,000
Senior Enterprise Platform Support Engineer – AI & Cloud
Senior Enterprise Platform Support Engineer – AI & Cloud

FIS • Bengaluru

On-site
INR 3,000,000 - 5,200,000
Competitive salary
Professional learning
Inclusive, diverse environment
+2
Analyst II, Production Support
Analyst II, Production Support

FIS Solutions (India) Private Limited - Pune • Pune District

On-site
INR 1,200,000 - 2,400,000
Senior DevOps Engineer(6+ years' experience in DevOps, AWS, Kubernetes and Docker)
Senior DevOps Engineer(6+ years' experience in DevOps, AWS, Kubernetes and Docker)

FIS • Pune District

On-site
INR 2,800,000 - 4,400,000
Competitive salary
Inclusive, diverse environment
Learning and development opportunities
Lead Site Reliability Engineer (SRE)
Lead Site Reliability Engineer (SRE)

FIS • Chennai District

Hybrid
INR 3,500,000 - 6,500,000
Collaboration culture
Competitive salary & benefits
Career growth opportunities
Senior Lead Java React Full Stack Engineer
Senior Lead Java React Full Stack Engineer

fis • Dadri

On-site
INR 4,200,000 - 7,000,000
GHMI/Hospitalization coverage for EMP-
Broad range of education and personal/
Senior Enterprise Platform Support Engineer AI & Cloud
Senior Enterprise Platform Support Engineer AI & Cloud

FIS • Bengaluru

On-site
INR 3,000,000 - 4,500,000
Java React Full Stack Engineer
Java React Full Stack Engineer

fis • Dadri

On-site
INR 1,500,000 - 2,600,000
GHMI/Hospitalization coverage
Professional education opportunities
Software Engineer Senior (.Net,Angular)
Software Engineer Senior (.Net,Angular)

fis • Pune District

On-site
INR 2,500,000 - 4,000,000