Site Reliability Engineering Senior Lead

FIS

Pune District

On-site

INR 2,500,000 - 4,000,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Engineering excellence culture
Continuous learning opportunities
Impactful projects on payment platforms

Job summary

FIS is hiring a Senior Lead Site Reliability Engineer in Pune to define and operate highly secure payment platforms for large-scale financial transactions. This senior role involves driving reliability outcomes across mission-critical systems and influencing architectural decisions.

Key responsibilities include leading incident responses, defining reliability architectures, and mentoring engineers. Ideal candidates will have deep expertise in software engineering, observability tools, and cloud platforms.

Qualifications

  • Deep expertise in distributed, API-driven systems.
  • Experience with observability and reliability tools.
  • Strong command of cloud infrastructure and automation.

Responsibilities

  • Own reliability outcomes for real-time distributed platforms.
  • Define reliability architecture and standards.
  • Lead responses to high-severity incidents.
  • Mentor engineers to improve operational maturity.

Skills

Software engineering expertise
Observability engineering
Cloud platforms (AWS, Azure, GCP)
Incident management
Automation skills (Python, Bash, Ansible)

Tools

Prometheus
Grafana
Datadog
ELK

Job description

Site Reliability Engineering Senior Lead – 10+ Yrs – UK Shift – Pune Location
About The Role

We are hiring a Senior Lead Site Reliability Engineer to define, build, and operate always‑on, low‑latency, and highly secure paymentplatforms that power large‑scale financial transactions.

This is a senior technical role, not a pure operations position. You will operate at the intersection of distributed systems engineering, cloud platforms, and reliability architecture, setting technical direction and driving reliability outcomes across mission‑critical, regulated systems in Payments and FinTech.

You will work across multiple teams and domains, influencing architecture, engineering practices, and operational maturity while remaining hands‑on with the most complex reliability challenges.

What You Will Do
  • Own and drive reliability outcomes at scale for real‑time, distributed payment and transaction processing platforms with strict SLAs, SLOs, and regulatory requirements.
  • Define reliability architecture and standards across services, platforms, and infrastructure, shaping how systems are designed, deployed, observed, and operated.
  • Design and evolve enterprise‑grade observability platforms (metrics, logs, traces, SLOs/SLIs) that provide actionable insights into system health, customer experience, and business impact.
  • Lead and coordinate response to high‑severity production incidents, acting as a technical authority during major events and driving deep root‑cause analysis and long‑term systemic fixes.
  • Set strategy and drive adoption of SRE best practices including error budgets, capacity modeling, resilience testing, graceful degradation, and operational readiness.
  • Architect automation and self‑service platforms that eliminate toil, reduce operational risk, and enable safe, frequent production releases across teams.
  • Partner with senior engineering, product, and platform leaders to influence architectural decisions, cloud migration strategy, disaster recovery posture, and long‑term platform evolution.
  • Mentor senior engineers and technical leads, raising the overall reliability and operational maturity of the organization.
What You Bring
  • Deep software engineering expertise with a proven track record of building and operating large‑scale, distributed, API‑driven systems in production.
  • Expertise in observability, alerting, and reliability engineering, using tools such as Prometheus, Grafana, Datadog, Splunk, ELK, or equivalent ecosystems.
  • Strong command of cloud platforms and open systems (AWS, Azure, or GCP), including infrastructure‑as‑code, platform automation, and cloud‑native design patterns.
  • Significant experience running mission‑critical systems in Payments, FinTech, Banking, or similarly regulated environments, where availability, correctness, and security are non‑negotiable.
  • Hands‑on experience across Linux (RHEL), Windows, databases (e.g., Oracle RDBMS), and complex enterprise stacks with strong system‑level troubleshooting skills.
  • Demonstrated leadership in incident management, post‑incident reviews, and continuous reliability improvement, with the ability to influence behavior and standards across teams.
  • Ability to operate effectively at Staff level scopesolving ambiguous problems, making trade‑offs, and driving alignment across multiple teams and stakeholders.
Added Advantage
  • Strong automation and scripting skills using Python, Bash, Ansible, or similar tools.
  • Experience building or scaling CI/CD platforms and release automation in high‑risk production environments.
  • Prior ownership of reliability strategy or platform initiatives spanning multiple teams or business units.
  • Experience modernizing legacy financial systems into cloud‑native or hybrid architectures with a focus on resilience and compliance.
Why Join Us
  • Work on high‑impact payment platforms operating at massive scale, where milliseconds and reliability directly affect real‑world commerce.
  • Play a Staff‑level role in defining reliability strategy for systems that cannot fail.
  • Join a culture that values engineering excellence, technical leadership, automation, and continuous learning.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Lead Site Reliability Engineer
Senior Lead Site Reliability Engineer

FIS • Pune District

On-site
INR 3,500,000 - 6,000,000
Site Reliability Engineer – Windows
Site Reliability Engineer – Windows

UBS • Maharashtra

On-site
INR 2,500,000 - 4,000,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Impronics Technologies • Gurugram District

On-site
INR 1,500,000 - 2,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Pune District

On-site
INR 1,200,000 - 1,800,000
Health insurance
Flexible working hours
Training opportunities
Site Reliability Engineer Lead
Site Reliability Engineer Lead

Hilabs • Pune District

On-site
INR 1,500,000 - 2,500,000
Site Reliability Engineer -2
Site Reliability Engineer -2

Groww • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

SMC Squared • Hyderabad

On-site
INR 3,500,000 - 6,000,000
Senior DevOps / Site Reliability Engineer (SRE)
Senior DevOps / Site Reliability Engineer (SRE)

Aura Recruitment Solutions • Bengaluru

On-site
INR 450,000 - 750,000
Lead SRE
Lead SRE

United States Digital Space LLC • Karnataka

On-site
INR 900,000 - 1,400,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

CellPoint Digital • Maharashtra

On-site
INR 3,000,000 - 6,000,000
Medical insurance with dependents