Senior Site Reliability Engineer, AI Platform

Thomson Reuters

Bengaluru

Hybrid

INR 800,000 - 1,100,000

Full time

19 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work model
Mental health days
Tuition reimbursement
Wellbeing programs

Job summary

Thomson Reuters seeks a Senior Site Reliability Engineer for the AI Platform to lead long-term infrastructure strategy and improve reliability across cloud-native services. You will design scalable systems, implement observability, enforce SLAs, and automate deployment pipelines while mentoring junior engineers.

Join a team focused on secure, scalable ML infra with hybrid work options and strong emphasis on DevOps best practices.

Qualifications

  • 5+ years of experience in cloud and infrastructure engineering.
  • Strong proficiency in Python and shell scripting.
  • Infrastructure-as-code experience with CloudFormation, Terraform, ARM or SAM.
  • Hands-on experience with cloud infrastructure.
  • Solid CI/CD tooling and deployment automation experience.
  • Familiarity with containerization (Docker, Kubernetes).

Responsibilities

  • Own and drive the long-term infrastructure strategy for AI Platform services.
  • Improve operational posture of cloud-native services including containers and serverless functions.
  • Instrument services with observability tooling (Datadog, CloudWatch, etc.).
  • Enforce SLAs and incident response processes; drive reliability improvements.
  • Collaborate with platform engineering to harden infrastructure against security findings.
  • Design and maintain CI/CD pipelines for AI Platform services and infra.
  • Automate toil: provisioning, security remediation, environment management, deployment ops.
  • Mentor junior team members and raise DevOps bar.

Skills

Cloud & infra engineering
Python
Shell scripting
IaC (CloudFormation/Terraform)
CI/CD
Docker
Kubernetes

Job description

About The Role
  • Own and drive the long-term infrastructure strategy for the AI Platform team — designing systems and standards that are built to scale, remain reliable under growth, and reduce operational burden over time.
  • Own and improve the operational posture of cloud-native services including containerized workloads, serverless functions, and managed ML infrastructure.
  • Instrument services with observability tooling (Datadog, CloudWatch, or equivalent) covering metrics, logs, and distributed tracing.
  • Establish and enforce SLAs and error budgets; drive incident response, post-mortems, and systemic reliability improvements.
  • Collaborate with platform engineering to harden infrastructure against security findings (Wiz, Snyk) and maintain compliance with organizational standards.
  • Design, build, and maintain CI/CD pipelines for AI Platform services, SDKs, and infrastructure, ensuring fast and reliable delivery across environments.
  • Automate toil — infrastructure provisioning, security remediation workflows, environment management, and deployment operations.
  • Mentor junior team members and raise the engineering bar across DevOps practices on the team.
About The Role
  • Own and drive the long-term infrastructure strategy for the AI Platform team — designing systems and standards that are built to scale, remain reliable under growth, and reduce operational burden over time.
  • Own and improve the operational posture of cloud-native services including containerized workloads, serverless functions, and managed ML infrastructure.
  • Instrument services with observability tooling (Datadog, CloudWatch, or equivalent) covering metrics, logs, and distributed tracing.
  • Establish and enforce SLAs and error budgets; drive incident response, post-mortems, and systemic reliability improvements.
  • Collaborate with platform engineering to harden infrastructure against security findings (Wiz, Snyk) and maintain compliance with organizational standards.
  • Design, build, and maintain CI/CD pipelines for AI Platform services, SDKs, and infrastructure, ensuring fast and reliable delivery across environments.
  • Automate toil — infrastructure provisioning, security remediation workflows, environment management, and deployment operations.
  • Mentor junior team members and raise the engineering bar across DevOps practices on the team.
In this opportunity as a Senior Site Reliability Engineer, AI Platform, you will:
  • Own and drive the long-term infrastructure strategy for the AI Platform team — designing systems and standards that are built to scale, remain reliable under growth, and reduce operational burden over time.
  • Own and improve the operational posture of cloud-native services including containerized workloads, serverless functions, and managed ML infrastructure.
  • Instrument services with observability tooling (Datadog, CloudWatch, or equivalent) covering metrics, logs, and distributed tracing.
  • Establish and enforce SLAs and error budgets; drive incident response, post-mortems, and systemic reliability improvements.
  • Collaborate with platform engineering to harden infrastructure against security findings (Wiz, Snyk) and maintain compliance with organizational standards.
  • Design, build, and maintain CI/CD pipelines for AI Platform services, SDKs, and infrastructure, ensuring fast and reliable delivery across environments.
  • Automate toil — infrastructure provisioning, security remediation workflows, environment management, and deployment operations.
  • Mentor junior team members and raise the engineering bar across DevOps practices on the team.
About You
  • Strong communication skills and the ability to work across engineering, security, and product teams and distill unclear requirements into clear engineering goals.
  • 5+ years of experience in cloud and infrastructure engineering roles.
  • Strong proficiency in Python and shell scripting, with hands-on experience building and deploying automation tools that eliminate toil and streamline DevOps workflows. Infrastructure-as-code experience with CloudFormation, Terraform, ARM or SAM is expected.
  • Deep hands-on experience with cloud infrastructure is expected.
  • Solid experience with CI/CD tooling and infrastructure deployment automation.
  • Familiar with containerization (Docker, Kubernetes) and cloud-native deployment patterns.
What’s in it For You?
  • Hybrid Work Model: We’ve adopted a flexible hybrid working environment (2-3 days a week in the office depending on the role) for our office-based roles while delivering a seamless experience that is digitally and physically connected.
  • Flexibility & Work-Life Balance: Flex My Way is a set of supportive workplace policies designed to help manage personal and professional responsibilities, whether caring for family, giving back to the community, or finding time to refresh and reset. This builds upon our flexible work arrangements, including work from anywhere for up to 8 weeks per year, empowering employees to achieve a better work-life balance.
  • Career Development and Growth: By fostering a culture of continuous learning and skill development, we prepare our talent to tackle tomorrow’s challenges and deliver real-world solutions. Our Grow My Way programming and skills-first approach ensures you have the tools and knowledge to grow, lead, and thrive in an AI-enabled future.
  • Industry Competitive Benefits: We offer comprehensive benefit plans to include flexible vacation, two company-wide Mental Health Days off, access to the Headspace app, retirement savings, tuition reimbursement, employee incentive programs, and resources for mental, physical, and financial wellbeing.
  • Culture: Globally recognized, award-winning reputation for inclusion and belonging, flexibility, work-life balance, and more. We live by our values: Obsess over our Customers, Compete to Win, Challenge (Y)our Thinking, Act Fast / Learn Fast, and Stronger Together.
  • Social Impact: Make an impact in your community with our Social Impact Institute. We offer employees two paid volunteer days off annually and opportunities to get involved with pro-bono consulting projects and Environmental, Social, and Governance (ESG) initiatives.
  • Making a Real-World Impact:We are one of the few companies globally that helps its customers pursue justice, truth, and transparency. Together, with the professionals and institutions we serve, we help uphold the rule of law, turn the wheels of commerce, catch bad actors, report the facts, and provide trusted, unbiased information to people all over the world.
About Us

Thomson Reuters informs the way forward by bringing together the trusted content and technology that people and organizations need to make the right decisions. We serve professionals across legal, tax, accounting, compliance, government, and media. Our products combine highly specialized software and insights to empower professionals with the data, intelligence, and solutions needed to make informed decisions, and to help institutions in their pursuit of justice, truth, and transparency. Reuters, part of Thomson Reuters, is a world leading provider of trusted journalism and news.

We are powered by the talents of 26,000 employees across more than 70 countries, where everyone has a chance to contribute and grow professionally in flexible work environments. At a time when objectivity, accuracy, fairness, and transparency are under attack, we consider it our duty to pursue them. Sound exciting? Join us and help shape the industries that move society forward.

As a global business, we rely on the unique backgrounds, perspectives, and experiences of all employees to deliver on our business goals. To ensure we can do that, we seek talented, qualified employees in all our operations around the world regardless of race, color, sex/gender, including pregnancy, gender identity and expression, national origin, religion, sexual orientation, disability, age, marital status, citizen status, veteran status or any other protected classification under applicable law. Thomson Reuters is proud to be an Equal Employment Opportunity Employer providing a drug-free workplace.

We also make reasonable accommodations for qualified individuals with disabilities and for sincerely held religious beliefs in accordance with applicable law. More information on requesting an accommodation here.

Learn more on how to protect yourself from fraudulent job postings here.

More information about Thomson Reuters can be found on thomsonreuters.com.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Jobtailor • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Hybrid Work Model
Flexible Work-Life Balance
Career Development and Growth
+3
Staff Software Engineer
Staff Software Engineer

Thomson Reuters • Hyderabad

Hybrid
INR 1,400,000 - 2,100,000
Hybrid Work Model
Flexible Work Arrangements
Career Development
+4
Automation and AI Solutions
Automation and AI Solutions

Jobtailor • Bengaluru

On-site
INR 1,400,000 - 2,000,000
Flexible vacation
Two mental health days
Tuition reimbursement
Staff Software Engineer
Staff Software Engineer

PowerToFly • Hyderabad

On-site
INR 3,000,000 - 5,500,000
Hybrid Work Model
Flex My Way
Career Growth
+2
Senior Software Engineer - AI II
Senior Software Engineer - AI II

PowerToFly • Bengaluru

On-site
INR 1,500,000 - 2,000,000
Hybrid Work Model
Flexible vacation
Career Development and Growth
Senior Site Reliability Engineer (SRE) I
Senior Site Reliability Engineer (SRE) I

Thomson Reuters • Bengaluru

Hybrid
INR 2,000,000 - 3,400,000
Hybrid work model
Flexible vacation
Mental Health Days
+3
Lead Security Engineer
Lead Security Engineer

Refinitiv • India

Hybrid
INR 2,800,000 - 5,200,000
Flexible vacation
Mental Health Days off
Headspace app
+4
Senior Site Reliability Engineer (SRE) I
Senior Site Reliability Engineer (SRE) I

Thomson Reuters • Hyderabad

Hybrid
INR 1,200,000 - 2,100,000
Flexible vacation
Mental Health Days
Headspace access
+2
Senior Data Scientist
Senior Data Scientist

Thomson Reuters • Bengaluru

Hybrid
INR 1,500,000 - 2,300,000
Hybrid Work Model
Flex My Way
Career Development
+1
Senior Software Engineer
Senior Software Engineer

Thomson Reuters • Bengaluru

Hybrid
INR 3,000,000 - 6,000,000