Senior Production Support Engineer

FIS

Pune District

On-site

INR 1,500,000 - 2,100,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Competitive salary and benefits
Professional development
Inclusive and diverse team

Job summary

FIS seeks a Senior Production Support Engineer to ensure stability, reliability, and performance of business-critical apps across cloud and hybrid environments. You will apply an SRE mindset to improve observability, reduce toil through automation, and resolve complex issues across infrastructure, apps, databases, and container platforms.

You will partner with Engineering, Infrastructure, Security, and Product teams to maintain service excellence, drive RCA actions, and develop runbooks,

Qualifications

  • Bachelor's degree in computer science or equivalent practical experience.
  • Proven Production Support or SRE experience.
  • Strong Azure and AWS infrastructure support experience.
  • Solid understanding of DNS, TCP/IP networking, routing, firewalls, and connectivity troubleshooting.
  • Hands-on with SQL Server administration and performance tuning.
  • Experience with AKS and Kubernetes environments.
  • Experience with GitHub and ArgoCD deployment pipelines.
  • Strong monitoring/observability with Dynatrace, Azure Monitor, Azure Log Analytics, and AWS CloudWatch.
  • Familiarity with distributed tracing and log analysis.
  • Knowledge of cloud-native architectures and microservices.
  • Experience conducting Root Cause Analysis and managing complex incidents.
  • Experience ensuring high availability of mission-critical systems.
  • Solid understanding of SRE principles, observability, automation, and reliability engineering.
  • Strong analytical and communication skills.

Responsibilities

  • Provide L2/L3 production support for business-critical applications.
  • Lead troubleshooting across apps, databases, cloud infra, networking, and Kubernetes.
  • Participate in incident response, major incident management, and on-call duties.
  • Conduct Root Cause Analysis and drive permanent fixes.
  • Develop and maintain runbooks, recovery procedures, and docs.
  • Monitor and improve availability, MTTR, service reliability, and performance metrics.
  • Support Azure and AWS environments, including networking, DNS, routing, firewalls, and VPNs.
  • Support AKS and Kubernetes-hosted apps, including workloads and ingress.
  • Support CI/CD pipelines with GitHub and ArgoCD; investigate deployment issues.
  • Design monitoring solutions using Dynatrace, Azure Monitor, Log Analytics, and CloudWatch.
  • Develop dashboards, alerts, SLIs, and SLOs; use logs, metrics, traces.
  • Identify automation opportunities to reduce toil and boost platform resilience.
  • Collaborate with Development and DevOps to improve production readiness.
  • Promote SRE best practices including blameless post-mortems and proactive service management.

Skills

Azure & AWS infra
DNS & networking
SQL Server administration
AKS & Kubernetes
GitHub & ArgoCD
Observability tooling
CI/CD pipelines
Distributed tracing
SRE fundamentals
Collaboration with DevOps

Education

Bachelor's degree in CS/IT/Engineering

Tools

Terraform / IaC
PowerShell / Bash / Python
Azure/AWS tooling

Job description

Senior Production Support Engineer
Intro

We are FIS. Our technology powers the worlds economy and our teams bring innovation to life. We champion diversity to deliver the best products and solutions for our colleagues, clients and communities. If youre ready to start learning, growing and making an impact with a career in fintech, wed like to know: Are you FIS

About the Role

As a Senior Production Support Engineer, you will ensure the stability, reliability, and performance of business-critical applications across cloud and hybrid environments. You will apply a Site Reliability Engineering (SRE) mindset to support production platforms, improve observability, reduce operational toil through automation, and resolve complex technical issues. Working across infrastructure, applications, databases, and container platforms, you will partner with Engineering, Infrastructure, Security, and Product teams to maintain service excellence. Success in this role is measured through service reliability, operational efficiency, incident resolution effectiveness, and continuous improvement outcomes.

What You Will Be Doing
  • Provide L2/L3 production support for business-critical applications and services.
  • Lead troubleshooting across applications, databases, cloud infrastructure, networking, and Kubernetes platforms.
  • Participate in incident response, major incident management, and on-call support activities.
  • Conduct detailed Root Cause Analysis (RCA) and drive permanent corrective actions.
  • Develop and maintain operational runbooks, recovery procedures, and support documentation.
  • Monitor and improve availability, MTTR, service reliability, and operational performance metrics.
  • Support and troubleshoot Azure and AWS environments, including networking connectivity, DNS, routing, firewalls, security groups, load balancers, hybrid connectivity, VPNs, and private network integrations.
  • Support AKS and Kubernetes-hosted applications, including workloads, namespaces, ingress controllers, and containerized services.
  • Support CI/CD deployment pipelines using GitHub and ArgoCD and investigate deployment and configuration issues.
  • Design and enhance monitoring and observability solutions using Dynatrace, Azure Monitor, Azure Log Analytics, and AWS CloudWatch.
  • Develop dashboards, alerts, SLIs, and SLOs while leveraging logs, metrics, traces, and distributed tracing data.
  • Identify automation opportunities, reduce manual operational effort, and contribute to platform reliability and resilience initiatives.
  • Collaborate with Development and DevOps teams to improve production readiness, platform stability, and operational maturity.
  • Promote SRE best practices including observability, automation, reliability engineering, blameless post-mortems, and proactive service management.
Required Qualifications
  • Bachelors degree in computer science, Information Technology, Engineering, or equivalent practical experience.
  • Proven experience in Production Support, Site Reliability Engineering, or a similar operational support function.
  • Strong Azure and AWS infrastructure support experience.
  • Strong understanding of DNS, TCP/IP networking, routing, firewalls, and connectivity troubleshooting.
  • Hands-on experience with Microsoft SQL Server administration, troubleshooting, performance tuning, locking, blocking, and query optimization.
  • Experience supporting AKS and Kubernetes environments.
  • Experience supporting GitHub and ArgoCD operational processes and deployment pipelines.
  • Strong monitoring and observability expertise with Dynatrace, Azure Monitor, Azure Log Analytics, and AWS CloudWatch.
  • Experience with distributed tracing, application performance monitoring, and log analysis.
  • Strong understanding of cloud-native architectures and microservices.
  • Demonstrated experience conducting Root Cause Analysis and managing complex production incidents.
  • Experience supporting high-availability, mission-critical production systems.
  • Strong knowledge of SRE principles, including observability, automation, reliability engineering, error budgets, Service Level Objectives (SLOs), and operational excellence.
  • Strong analytical, problem-solving, stakeholder management, and communication skills.
Preferred Qualifications
  • Experience supporting financial services or highly regulated environments.
  • Knowledge of Infrastructure as Code tools including Terraform, ARM, Bicep, or CloudFormation.
  • Experience with PowerShell, Bash, Python, or automation scripting.
  • Experience implementing operational dashboards and automated remediation solutions.
  • Understanding of DevOps practices and CI/CD methodologies.
  • Industry certifications in Azure, AWS, Kubernetes, or Site Reliability Engineering disciplines.
  • Ability to troubleshoot complex issues across multiple technology layers.
  • Customer-focused mindset with a passion for automation, reliability, and continuous improvement.
What We Offer you

At FIS, we are as committed to growing our employees careers as our own business. We offer:

  • Opportunities to innovate in fintech
  • Inclusive and diverse team atmosphere
  • Professional and personal development
  • Resources to contribute to your community
  • Competitive salary and benefits
Privacy Statement

FIS is committed to protecting the privacy and security of all personal information that we process in order to provide services to our clients. For specific information on how FIS protects personal information online, please see the .

Sourcing Model

#pridepass

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE Production Support (Rotational Shift)
SRE Production Support (Rotational Shift)

FIS India Bangalore • India

On-site
INR 1,200,000 - 2,200,000
Production Support Analyst I(AWS and Linux/UNIX administration)
Production Support Analyst I(AWS and Linux/UNIX administration)

FIS Solutions (India) Private Limited - Pune • Pune District

On-site
INR 1,500,000 - 2,300,000
Senior Enterprise Platform Support Engineer – AI & Cloud
Senior Enterprise Platform Support Engineer – AI & Cloud

FIS • Bengaluru

On-site
INR 3,000,000 - 5,200,000
Competitive salary
Professional learning
Inclusive, diverse environment
+2
Senior Enterprise Platform Support Engineer – AI & Cloud
Senior Enterprise Platform Support Engineer – AI & Cloud

FIS Solutions (India) Private Limited - Bangalore • Pune District

On-site
INR 4,000,000 - 8,000,000
Senior Enterprise Platform Support Engineer AI & Cloud
Senior Enterprise Platform Support Engineer AI & Cloud

FIS • Bengaluru

On-site
INR 3,000,000 - 4,500,000
Site Reliability Engineer
Site Reliability Engineer

fis • Pune District

On-site
INR 1,200,000 - 1,800,000
Private medical cover
Dental cover
Travel insurance
+1
Production Support Analyst I(AWS and Linux/UNIX administration)
Production Support Analyst I(AWS and Linux/UNIX administration)

FIS • Pune District

On-site
INR 1,200,000 - 1,800,000
Senior DevOps Analyst – Cloud and AI Automation
Senior DevOps Analyst – Cloud and AI Automation

FIS Solutions (India) Private Limited - Pune • Pune District

On-site
INR 4,200,000 - 6,500,000
Competitive compensation and benefits
Learning and development opportunities
Inclusive, diverse work environment
+1
Site Reliability Engineer (SRE) – Cloud Platform
Site Reliability Engineer (SRE) – Cloud Platform

FIS Solutions (India) Private Limited - Pune • Pune District

On-site
INR 1,200,000 - 2,000,000
Long-term ownership
Global collaboration with multi-region
Learning & certification support
+1
Support Engineer Enterprise Platforms
Support Engineer Enterprise Platforms

FIS • Bengaluru

On-site
INR 3,500,000 - 5,500,000