Site Reliability Engineer

Thales

Dadri

On-site

INR 1,200,000 - 1,800,000

Full time

8 days ago
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Thales in Noida, India, is seeking a reliability and operations professional to ensure high availability and performance of cloud services. You will lead incident response, perform root-cause analyses, and drive long-term fixes across Dev and Ops teams.

The role involves building IaC and automation with Terraform and GitLab CI/CD, supporting AWS/GCP deployments, and enhancing observability with Datadog and Splunk. Collaboration with product owners is essential.

Qualifications

  • Experience with Kubernetes, AWS and/or GCP in production environments.
  • Strong knowledge of CI/CD pipelines and Infrastructure as Code (GitLab CI, Terraform).
  • Familiarity with SRE principles (SLIs, SLOs, error budgets) and incident management.

Responsibilities

  • Ensure high availability, performance, and scalability of Thales PAY Digital cloud services in line with SLAs/SLOs.
  • Respond to production incidents, lead troubleshooting, and coordinate recovery actions.
  • Conduct post-incident analyses and implement long-term corrective actions.
  • Participate in on-call and incident escalation processes.
  • Design, develop, and maintain IaC and automation solutions (Terraform, scripting).
  • Support and validate cloud deployments of Thales PAY products (AWS / GCP / Kubernetes).
  • Provide technical input for changes and release readiness.

Skills

Kubernetes
AWS/GCP
CI/CD pipelines
Linux
Observability tools
SRE principles
Shell/Python scripting
Agile & DevOps

Tools

Terraform
GitLab CI/CD
Datadog
Splunk

Job description

Location: Noida, India

Thales is a global technology leader trusted by governments, institutions, and enterprises to tackle their most demanding challenges. From quantum applications and artificial intelligence to cybersecurity and 6G innovation, our solutions empower critical decisions rooted in human intelligence. Operating at the forefront of aerospace and space, cybersecurity and digital identity, we’re driven by a mission to build a future we can all trust. Present in India since 1953, Thales is headquartered in Noida and has other operational offices and sites spread across Delhi, Gurugram, Bengaluru and Mumbai, among others. Over 2200 employees are working with Thales and its joint ventures in India. Since the beginning, Thales has been playing an essential role in India’s growth story by sharing its technologies and expertise in Defence, Aerospace and Cyber & Digital sectors. Thales has two engineering competence centres in India - one in Noida focused on Cyber & Digital business, while the one in Bengaluru focuses on hardware, software and systems engineering capabilities for both the civil and defence sectors, serving global needs. The Group has also established an MRO (Maintenance, Repair & Overhaul) facility in Gurugram to provide comprehensive avionics maintenance and repair services to Indian airlines and support the growth of the local aviation industry.

Missions And Responsibilities

Reliability & Operations

  • Ensure high availability, performance, and scalability of Thales PAY Digital cloud services in line with customer SLAs/SLOs
  • Respond to production incidents, lead troubleshooting efforts, and coordinate recovery actions
  • Conduct post-incident analyses (RCA/postmortems) and implement long-term corrective and preventive actions
  • Participate in on-call and incident escalation processes

Automation & Infrastructure

  • Design, develop, and maintain Infrastructure as Code (IaC) and automation solutions (Terraform, GitLab CI/CD, scripting)
  • Support and validate cloud deployments of Thales PAY products (AWS / GCP / Kubernetes)

Change & Release Management

  • Provide technical expertise and risk assessment during Change Advisory Board (CAB) activities, especially for high-risk changes impacting customer SLAs
  • Ensure proper change planning, validation, and rollback strategies
  • Contribute to release readiness and operational acceptance

Observability & Performance

  • Implement and maintain monitoring, alerting, and observability across platforms (metrics, logs, traces)
  • Continuously improve system monitoring, dashboards, and alert quality using tools such as Datadog and Splunk
  • Perform capacity planning, performance tuning, and continuous technological watch

Collaboration & Documentation

  • Work closely with Product Owners and development squads to anticipate operational needs
  • Provide technical input for new services and service evolutions
Skills Requirements (Domain/technical/soft)

Technical Skills

  • Experience in systems, networking, and security operations
  • Hands-on experience with Kubernetes, AWS, and/or GCP in production environments
  • Strong experience with CI/CD pipelines and Infrastructure as Code (GitLab CI, Terraform)
  • Proficiency in Linux, TCP/IP, HTTP/HTTPS, and distributed systems
  • Strong knowledge of observability and monitoring tools (Datadog, Splunk, logs/metrics/traces)
  • Solid scripting skills (Shell and/or Python)

Methodologies & Practices

  • Knowledge of SRE principles (SLIs, SLOs, error budgets, automation, incident management)
  • Familiarity with Agile and DevOps ways of working
  • Experience in service delivery, operational readiness, and production support

Soft Skills

  • Strong analytical and troubleshooting skills
  • Ability to diagnose and resolve complex, high pressure production issues
  • Clear communication with both technical and non-technical stakeholders
  • Ownership mindset and customer-oriented approach

At Thales, we’re committed to fostering a workplace where respect, trust, collaboration, and passion drive everything we do. Here, you’ll feel empowered to bring your best self, thrive in a supportive culture, and love the work you do. Join us, and be part of a team reimagining technology to create solutions that truly make a difference – for a safer, greener, and more inclusive world.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

TechOps Engineer III - Noida (24/7 shift)
TechOps Engineer III - Noida (24/7 shift)

Thales • Dadri

Hybrid
INR 1,200,000 - 1,800,000
Java Developer
Java Developer

Thales • Dadri

On-site
INR 500,000 - 800,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Thales Group • Bengaluru

On-site
INR 1,300,000 - 1,800,000
TechOps Engineer III
TechOps Engineer III

Thales • Dadri

On-site
INR 1,200,000 - 2,400,000
TechOps Engineer III
TechOps Engineer III

Thales Group • India

On-site
INR 2,600,000 - 4,200,000
Full Stack Software Engineer
Full Stack Software Engineer

Thales • Bengaluru

On-site
INR 1,400,000 - 2,000,000
Escalation Engineer – Client Services
Escalation Engineer – Client Services

Thales • Dadri

On-site
INR 700,000 - 1,200,000
Supportive work environment
Opportunities for professional growth
Commitment to diversity and inclusion
Technical Lead - Java Backend
Technical Lead - Java Backend

Thales • Bengaluru Urban

On-site
INR 2,400,000 - 3,600,000
Engineer (Software Solutions)
Engineer (Software Solutions)

Thales Group • Bengaluru

On-site
INR 1,500,000 - 1,900,000
Sr Software Engineer (Gloang Developer)
Sr Software Engineer (Gloang Developer)

thales • Dadri

On-site
INR 1,800,000 - 3,000,000