Site Reliability Engineer – Windows

UBS

Maharashtra

On-site

INR 2,500,000 - 4,000,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

UBS in Pune is hiring an experienced Site Reliability Engineer to join our mission-critical technology team. You will design, implement, and operate reliable, scalable systems in a regulated financial environment, embedding reliability into the software delivery lifecycle.

You will partner with global Windows OS and middleware teams, define SLIs/SLOs/SLAs, automate toil away, manage incidents, and integrate with Prometheus, Grafana, ELK, and Datadog to provide end-to-end visibility across time

Qualifications

  • Proven expertise in Site Reliability Engineering with a software/infra background.
  • Hands-on experience with cloud platforms (Azure DevOps) and Windows Server 2019+.
  • Strong networking and storage knowledge (NFS, SAN, NAS).
  • Proficiency in scripting and automation (Python, PowerShell).
  • Ability to define and manage SLIs/SLOs/SLAs and reduce TOIL.
  • Familiarity with observability tools (Prometheus, Grafana, ELK, Datadog).

Responsibilities

  • Design and maintain highly available, fault-tolerant systems in a financial environment.
  • Define and monitor SLIs, SLOs, and SLAs for reliability.
  • Lead incident response, post-mortems and root cause analysis.
  • Collaborate with development teams to embed reliability in SDLC.
  • Automate repetitive tasks to reduce TOIL.
  • Integrate with observability platforms for end-to-end visibility.

Skills

SRE expertise
Cloud platforms
Networking fundamentals
Automation scripting
Observability tooling
Incident response
Collaboration

Tools

Azure DevOps
Terraform
GitLab
Prometheus
Grafana
ELK
Datadog
Kubernetes
AKS

Job description

Your role: We are seeking a highly experienced Site Reliability Engineer (SRE) to join our technology team in a mission‑critical financial environment. This role is ideal for someone who has a proven track record of building and operating reliable, scalable systems in regulated industries such as banking or financial services.

Job Type: Full Time

Job Reference #: 337481BR

City: Pune

Your team: You’ll be joining the Windows team in Operating Systems and Middleware (OSM) crew, a globally distributed team that supports critical infrastructure across time zones using a follow‑the‑sun support model. Based in Pune/Hyderabad, this role is embedded in a collaborative Agile environment where engineers are empowered to take ownership, innovate, and continuously improve.

Key Responsibilities
  • Design, implement, and maintain highly available and fault‑tolerant systems in a financial environment.
  • Define and monitor Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs) to ensure system reliability and customer satisfaction.
  • Identify, measure, and reduce TOIL, proactively eliminating repetitive manual tasks through automation.
  • Lead incident response, post‑mortems, and root cause analysis for production issues.
  • Collaborate with development teams to embed reliability into the software development lifecycle.
  • Integrate with observability platforms (Prometheus, Grafana, ELK, Datadog) to ensure end‑to‑end visibility of systems and services.
Essential Experience & Skills
  • Proven expertise in Site Reliability Engineering, with a background in software engineering, infrastructure, or operations.
  • Hands‑on experience with cloud platforms (e.g., Azure DevOps), operating systems (e.g., Windows Server 2019+), and networking fundamentals.
  • Solid understanding of networking and storage technologies (e.g., NFS, SAN, NAS).
  • Strong working knowledge of authentication and naming services (e.g., DNS, LDAP, Kerberos, Centrify).
  • Proficiency in scripting and automation (Python, PowerShell).
  • Experience with Azure ADO pipelines and Azure control plane.
  • Experience with infrastructure‑as‑code tools (Terraform, GitLab).
  • Demonstrated ability to define and manage SLIs, SLOs, SLAs, and systematically reduce TOIL.
  • Ability to integrate with observability platforms for system visibility.
  • Knowledge of Microsoft Failover clustering and DFS architecture.
  • Metrics‑ and automation‑driven mindset focused on measurable reliability.
  • Calm under pressure during incidents and outages, with a structured approach to incident response and post‑mortems.
  • Strong collaboration and communication skills, working across engineering and business teams.
  • Proactive, ownership‑driven attitude, continuously seeking ways to improve systems and processes.
Desirable Additions
  • Experience with chaos engineering, resilience testing, or disaster recovery planning.
  • Familiarity with financial transaction systems, real‑time data pipelines, or core banking platforms.
  • Understanding of CI/CD pipelines, containerization (AKS), and orchestration (Kubernetes).
  • Curiosity to explore how AI can improve workflows, validated against policy, risk, and ethics.

UBS is an Equal Opportunity Employer. We respect and seek to empower each individual and support the diverse cultures, perspectives, skills and experiences within our workforce. We’re committed to disability inclusion and if you need reasonable accommodation or adjustments throughout our recruitment process, you can contact us.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Persistent Systems Limited • Pune District

On-site
INR 2,500,000 - 4,200,000
Competitive salary
Benefits package
Talent development
+4
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Pune District

On-site
INR 1,200,000 - 1,800,000
Health insurance
Flexible working hours
Training opportunities
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Persistent Systems Limited • Pune District

On-site
INR 1,400,000 - 2,200,000
Hybrid work
Long Service awards
Company-sponsored education
Site Reliability Engineer (Azure Preferred)
Site Reliability Engineer (Azure Preferred)

FIS Solutions (India) Private Limited - Pune • Pune District

On-site
INR 2,500,000 - 6,000,000
Site Reliability Engineer (Azure Preferred)
Site Reliability Engineer (Azure Preferred)

fis • Pune District

On-site
INR 1,200,000 - 1,800,000
Competitive salary
Flexible work environment
Career growth opportunities
Lead SRE
Lead SRE

UST • Bengaluru

On-site
INR 3,000,000 - 5,500,000
Associate - SRE - Platform Engineering
Associate - SRE - Platform Engineering

Jefferies Financial Group Inc. • Pune District

On-site
INR 1,600,000 - 2,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

BayOne Solutions • Hyderabad

On-site
INR 3,000,000 - 6,000,000
Site Reliability Engineer -2
Site Reliability Engineer -2

Groww • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Reliability Engineer
Reliability Engineer

Us Bank • Chennai District

Hybrid
INR 1,500,000 - 2,100,000