Senior Software Engineer- Site Reliability

S&P Global

Hyderabad

On-site

INR 2,500,000 - 4,200,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Health care coverage
Flexible time off
Continuous learning
Retirement planning

Job summary

Kensho, a Kensho unit within S&P Global, seeks a Senior Site Reliability Engineer to own reliability and security of internal and customer-facing services. You will lead design, automation, and on-call incident response in a 24/7 environment.

You'll work with AWS/EKS, Terraform, Kubernetes, and Python to build scalable systems, improve observability, and partner with InfoSec to maintain a strong security posture while driving cost efficiency and platform resiliency.

Qualifications

  • 6+ years in SRE, DevOps, or Infrastructure roles.
  • Strong Python-based automation for reliability.
  • Experience with large-scale distributed systems in production.
  • Deep AWS experience including IAM and networking.
  • Hands-on Kubernetes (EKS) operations and troubleshooting.
  • CI/CD and infrastructure as code expertise.
  • Knowledge of databases and performance under load.

Responsibilities

  • Own and operate production services for critical financial applications with focus on availability.
  • Design, build, and manage AWS infrastructure across environments.
  • Provision infrastructure using Terraform with automation-first mindset.
  • Deploy, scale, and troubleshoot Kubernetes workloads and clusters.
  • Develop automation tooling to reduce toil and prevent incidents.
  • Monitor health via metrics/logs/alerts; refine runbooks and dashboards.
  • Lead on-call, drive incident response, and perform root-cause analyses.
  • Collaborate with InfoSec and security teams to maintain a strong posture.

Skills

Python automation
AWS
Kubernetes (EKS)
Terraform
On-call incident management

Tools

Terraform
Python
Kubernetes
AWS
PostgreSQL
Kafka
Linux (Ubuntu)

Job description

Job Description:
About The Role:

Grade Level (for internal use): 03

Who We Are:

Kensho is S&P Global's hub for AI innovation and transformation. With expertise in Machine Learning and data discovery, we develop and deploy novel solutions for S&P Global and its customers worldwide. Our solutions help businesses harness the power of data and Artificial Intelligence to innovate and drive progress. Kensho's solutions and research focus on Generative AI, LLM Agents, speech recognition, entity linking, document extraction, text classification, natural language processing, and more.

At Kensho, we hire talented people and give them the autonomy and support needed to build amazing technology and products. We collaborate using our teammates' diverse perspectives to solve hard problems. Our communication with one another is open, honest, and efficient. We dedicate time and resources to explore new ideas, but always rooted in engineering best practices. As a result, we can innovate rapidly to produce technology that is scalable, robust, and useful.

About The Role:

As a Senior Site Reliability Engineer (SRE) at Kensho, you will be a hands‑on technologist who combines strong infrastructure expertise with solid software engineering skills Python first. You will be responsible for ensuring the reliability, scalability, and security of both business‑critical internal systems and external, customer facing services.

You will work closely with Infrastructure, Application, and Security teams to design resilient systems, automate operations, and continuously improve platform stability. This role requires deep ownership of production systems, strong troubleshooting skills across infrastructure, Container orchestration systems, networking, and applications, and comfort operating in a 24/7 on call environment.

What You’ll Do:
  • Own and operate production services supporting critical financial applications with a strong focus on availability, performance, and reliability
  • Design, build, and manage AWS infrastructure, including EKS-based clusters, across lower and production environments
  • Provision and manage infrastructure using Terraform (Infrastructure as Code) with a strong automation first mindset
  • Deploy, scale, and troubleshoot applications running on like Kubernetes, including cluster creation, upgrades, and lifecycle management
  • Build and maintain automation frameworks and tooling Python based to reduce operational toil and prevent recurring incidents
  • Monitor system health using metrics, logs, and alerts; continuously tune alerts, dashboards, and runbooks
  • Troubleshoot complex issues spanning clusters, networking, certificates, deployments, and application behavior
  • Manage certificate lifecycle and expiration, ensuring secure and uninterrupted service operation
  • Collaborate with InfoSec, Vulnerability Management, and Network Security teams (e.g., Zscaler) to maintain a strong security posture
  • Collaborate with L1/L2 teams, helping them understand infrastructure and operational best practices
  • Participate in on call and lead incident response, drive root cause analysis, and ensure effective post‑incident remediation and learnings
  • Identify architectural anti‑patterns and drive improvements by reviewing new services for production readiness, resiliency, and secure design prior to release
  • Establish and enforce production readiness standards, including deployment strategies, rollback plans, and observability requirements
  • Optimize infrastructure cost and resource utilization without compromising reliability and performance
What We Look For:
  • 6+ years of experience in SRE, DevOps, Platform, or Infrastructure Engineering roles
  • Strong software engineering background, with hands‑on Python development used for automation, tooling, and system reliability
  • Experience building or supporting scalable, distributed systems in production
  • Deep experience with AWS cloud environments, including AWS, IAM, networking, and access controls
  • Strong hands‑on expertise with similar tools like Kubernetes (EKS preferred): cluster creation, deployments, scaling, and troubleshooting
  • Solid understanding of networking fundamentals (VPCs, routing, DNS, load balancing, security groups)
  • Experience with CI/CD pipelines, deployment tools, and infrastructure automation
  • Working knowledge of databases and query optimization, and understanding how applications behave under load
  • Familiarity with similar tools like Kafka or other messaging systems
  • Comfortable conducting code reviews and participating in coding focused interviews
  • Strong operational mindset with experience in incident management and on‑call rotations
  • Clear communicator and collaborative teammate who values documentation and knowledge sharing
How To Really Get Our Attention:
  • Demonstrated ownership of large‑scale, production systems
  • Strong examples of Python based automation or internal tooling
  • Contributions to open‑source projects, infrastructure platforms, or reliability tooling
  • Experience working closely with security and compliance teams in regulated environments
Technologies We Like:
  • AWS, Amazon EKS, Terraform, Jsonnet
  • Similar tools like Kubernetes, Helm, CI/CD tooling
  • Python (automation, tooling, reliability engineering)
  • Prometheus, Grafana, logging and monitoring platforms
  • PostgreSQL and other production databases
  • Similar tools like Kafka or event driven systems
  • Linux (Ubuntu or similar)
What’s In It For You?
Our Mission:

Advancing Essential Intelligence.

Our People:

We're more than 35,000 strong worldwide- so we're able to understand nuances while having a broad perspective. Our team is driven by curiosity and a shared belief that Essential Intelligence can help build a more prosperous future for us all. From finding new ways to measure sustainability to analyzing energy transition across the supply chain to building workflow solutions that make it easy to tap into insight and apply it. We are changing the way people see things and empowering them to make an impact on the world we live in. We're committed to a more equitable future and to helping our customers find new, sustainable ways of doing business. Join us and help create the critical insights that truly make a difference.

Our Values:
Integrity, Discovery, Partnership

Throughout our history, the world's leading organizations have relied on us for the Essential Intelligence they need to make confident decisions about the road ahead. We start with a foundation of integrity in all we do, bring a spirit of discovery to our work, and collaborate in close partnership with each other and our customers to achieve shared goals.

Benefits:

We take care of you, so you can take care of business. We care about our people. That’s why we provide everything you—and your career—need to thrive at S&P Global.

Our Benefits Include:
  • Health & Wellness: Health care coverage designed for the mind and body.
  • Flexible Downtime: Generous time off helps keep you energized for your time on.
  • Continuous Learning: Access a wealth of resources to grow your career and learn valuable new skills.
  • Invest in Your Future: Secure your financial future through competitive pay, retirement planning, a continuing education program with a company-matched student loan contribution, and financial wellness programs.
  • Family Friendly Perks: It's not just about you. S&P Global has perks for your partners and little ones, too, with some best-in class benefits for families.
  • Beyond the Basics: From retail discounts to referral incentive awards- small perks can make a big difference.

For more information on benefits by country visit: https://spgbenefits.com/benefit-summaries

Global Hiring And Opportunity At S&P Global:

At S&P Global, we are committed to fostering a connected and engaged workplace where all individuals have access to opportunities based on their skills, experience, and contributions. Our hiring practices emphasize fairness, transparency, and merit, ensuring that we attract and retain top talent. By valuing different perspectives and promoting a culture of respect and collaboration, we drive innovation and power global markets.

Equal Opportunity Employer

S&P Global is an equal opportunity employer and all qualified candidates will receive consideration for employment without regard to race/ethnicity, color, religion, sex, sexual orientation, gender identity, national origin, age, disability, marital status, military veteran status, unemployment status, or any other status protected by law. Only electronic job submissions will be considered for employment.

If you need an accommodation during the application process due to a disability, please send an email to: EEO.Compliance@spglobal.comand your request will be forwarded to the appropriate person.

US Candidates Only:

Know Your Rights: Workplace discrimination is illegal

20 - Professional (EEO-2 Job Categories-United States of America), BSMGMT203 - Entry Professional (EEO Job Group)

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead I, Software Development/Engineering
Lead I, Software Development/Engineering

S&P Global • Ahmedabad District

On-site
INR 3,500,000 - 6,500,000
Health care coverage
Flexible downtime
Continuous learning
+2
Software Development Engineer in Test (SDET) – Java
Software Development Engineer in Test (SDET) – Java

OSTTRA • Haryana

On-site
INR 1,200,000 - 2,400,000
Health & Wellness
Flexible Downtime
Continuous Learning
+2
Cloud Security Specialist
Cloud Security Specialist

S&P Global, Inc. • Hyderabad

On-site
INR 1,800,000 - 3,200,000
Health & Wellness benefits
Flexible downtime
Continuous learning
+3
Machine Learning Engineer
Machine Learning Engineer

S&P Global • Hyderabad

On-site
INR 1,800,000 - 2,800,000
Health & Wellness
Flexible Downtime
Continuous Learning
+3
Machine Learning Engineer
Machine Learning Engineer

S&P Global • Bengaluru

On-site
INR 1,500,000 - 2,800,000
Senior Data Scientist
Senior Data Scientist

S&P Global, Inc. • Gurgaon

On-site
INR 1,800,000 - 3,000,000
Health care
Flexible time off
Continual learning
+2
Senior Software Engineer- Site Reliability
Senior Software Engineer- Site Reliability

S&P Global • Bengaluru

On-site
INR 6,535,000 - 10,271,000
Health & Wellness
Flexible Downtime
Continuous Learning
+3
Senior Machine Learning Engineer
Senior Machine Learning Engineer

S&P Global • Hyderabad

On-site
INR 2,500,000 - 5,000,000
Health & Wellness
Flexible Downtime
Continuous Learning
+3
Senior Software Engineer- Site Reliability
Senior Software Engineer- Site Reliability

S&P Global • Hyderabad

On-site
INR 1,500,000 - 2,500,000
Health care coverage
Generous time off
Access to career resources
+2
Geopolitical Analyst/OSINT Analyst
Geopolitical Analyst/OSINT Analyst

S&P Global, Inc. • Bengaluru

On-site
INR 600,000 - 900,000
Health & Wellness
Flexible Downtime
Continuous Learning
+3