Lead II - DevOps Engineering SRE DevOps

UST

Thiruvananthapuram

On-site

INR 3,000,000 - 4,500,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

UST is seeking a hands-on DevOps SRE Engineer to support platform engineering for hosting applications and APIs in AWS and Kubernetes environments. You will operate, troubleshoot, and optimize platforms in a multi-vendor ecosystem, coordinating with Tier 2 teams and cross-functional groups.

The role emphasizes production incident management, automation via Terraform, and strong communication to drive platform reliability and transformation initiatives.

Qualifications

  • Hands-on DevOps/SRE experience in production environments.
  • Strong AWS production infra management and troubleshooting.
  • Kubernetes hands-on experience in hosting platforms.
  • Experience with CI/CD pipelines, preferably GitLab CI/CD.
  • Hands-on Terraform and IaC experience.
  • Proficient in production troubleshooting and incident management.
  • Linux/Unix troubleshooting expertise.
  • Experience supporting application and API hosting platforms.
  • Platform reliability and observability improvement focus.
  • Strong communication and stakeholder management.

Responsibilities

  • Maintain, troubleshoot, and enhance CI/CD pipelines (GitLab preferred).
  • Manage and support AWS infrastructure and cloud-native services.
  • Deploy, operate, and troubleshoot Kubernetes workloads and platforms.
  • Perform production incident triage, debugging, and RCA.
  • Utilize logs, metrics, monitoring tools (e.g., Splunk) for troubleshooting.
  • Develop and enhance Terraform-based infrastructure and automation.
  • Support platform modernization and engineering transformation initiatives.
  • Drive improvements in platform stability, reliability, observability, and performance.
  • Collaborate with GCC team and UST BFF leads across vendors.
  • Lead technical discussions with Tier 2 support and cross-functional teams.
  • Identify recurring issues and implement sustainable platform solutions.

Skills

DevOps / SRE
AWS
Kubernetes
GitLab CI/CD
Terraform
IaC
Incident management
Linux
Communication
Multi-vendor coordination

Tools

GitLab
Terraform
Helm
Docker
Prometheus
Grafana
CloudWatch
Splunk

Job description

Role Description

Experience: 8-15 Years

Job Location

Chennai, Bangalore, Hyderabad, Kochi, Trivandrum, Noida, Pune

DevOps SRE Engineer / Platform Engineer Role Overview

We are looking for highly hands‑on DevOps SRE Engineers with strong Platform Engineering experience supporting application and API hosting platforms in AWS and Kubernetes environments.

The ideal candidate should have strong recent hands‑on technical expertise in production environments, with the ability to operate, troubleshoot, optimize, and enhance platforms within a multi‑vendor ecosystem involving multiple teams.

This is a hands‑on individual contributor role and not a people‑management position. Strong communication skills are essential, along with the ability to lead technical discussions with Tier 2 teams, stakeholders, and cross‑functional engineering groups.

Key Responsibilities
  • Maintain, troubleshoot, and enhance CI/CD pipelines, with GitLab preferred.
  • Manage and support AWS infrastructure and cloud‑native services.
  • Deploy, operate, and troubleshoot Kubernetes workloads and platforms.
  • Perform production incident triage, debugging, and root cause analysis.
  • Use logs, metrics, monitoring tools, and Splunk for effective production troubleshooting.
  • Develop and enhance Terraform‑based infrastructure and automation.
  • Support repository restructuring, platform modernization, and engineering transformation initiatives.
  • Drive improvements in platform stability, reliability, observability, scalability, and performance.
  • Support platforms that host and enable application and API services.
  • Collaborate closely with the GCC team and UST BFF leads.
  • Work effectively across a multi‑vendor engineering ecosystem and coordinate technical resolution across teams.
  • Drive technical discussions, identify recurring platform issues, and implement sustainable solutions.Contribute to continuous improvement of platform engineering practices, automation, and operational processes.
Must‑Have Skills
  • Strong hands‑on experience in DevOps / Site Reliability Engineering (SRE).
  • Strong AWS experience, including production infrastructure management and troubleshooting.
  • Strong hands‑on Kubernetes experience.
  • Strong experience with CI/CD pipelines, preferably GitLab CI/CD.
  • Hands‑on experience with Terraform and Infrastructure as Code (IaC).
  • Strong production troubleshooting and incident management experience.Experience with Splunk, logs, metrics, and monitoring/observability tools.
  • Strong understanding of Linux/Unix environments and troubleshooting.
  • Experience supporting application and API hosting platforms.
  • Experience with platform reliability, availability, performance, and observability improvements.
  • Strong debugging and root cause analysis (RCA) capabilities.
  • Ability to work independently as a hands‑on technical contributor.
  • Strong communication and stakeholder‑management skills.
  • Ability to drive technical discussions with Tier 2 support, engineering teams, stakeholders, and cross‑functional teams.
  • Experience working in a multi‑vendor / distributed engineering environment.
Good‑to‑Have Skills
  • Experience with AWS EKS or other managed Kubernetes services.
  • Experience with GitLab administration or advanced GitLab CI/CD.
  • Advanced Terraform modules, state management, and automation experience.
  • Experience with Helm and Kubernetes deployment automation.
  • Experience with observability platforms such as Prometheus, Grafana, CloudWatch, or similar tools.
  • Experience with Docker/containerization.
  • Experience with scripting/automation using Python, Shell, or Bash.
  • Knowledge of API platforms, microservices, and cloud‑native architectures.
  • Experience with platform modernization and repository restructuring.
  • Experience implementing SRE practices, SLIs/SLOs, error budgets, and reliability engineering principles.
  • Experience with security, IAM, networking, and cost optimization in AWS.
  • Experience coordinating technical initiatives across multiple vendors and engineering teams.
Experience Range

6–12 years of overall experience in DevOps, SRE, Cloud Engineering, Platform Engineering, or related roles.

Preferred: 4+ years of strong hands‑on experience with AWS and Kubernetes in production environments.

Candidates with extensive recent hands‑on expertise and strong production troubleshooting experience will be preferred over candidates with primarily managerial or coordination experience.

Preferred Candidate Profile

The ideal candidate is a strong hands‑on engineer who can operate production platforms, troubleshoot complex incidents, automate infrastructure, improve reliability, and work across multiple engineering/vendor teams. The role requires someone who can not only execute technically but also confidently drive technical conversations and influence stakeholders toward effective platform solutions.

Skills

Site Reliability Engineering, Splunk, Terraform, Kubernetes

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Hilabs • Bengaluru

On-site
INR 2,200,000 - 3,500,000
Sr. Lead DevOps Engineer
Sr. Lead DevOps Engineer

Learningmate Solutions • Mumbai

On-site
INR 4,000,000 - 6,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Falabella India • Bengaluru

On-site
INR 4,000,000 - 7,000,000
DevOps & Site Reliability Engineer (SRE)
DevOps & Site Reliability Engineer (SRE)

Kiya.ai • Chennai District

On-site
INR 3,000,000 - 5,000,000
Senior Site Reliability Engineer (SRE) Engineer
Senior Site Reliability Engineer (SRE) Engineer

Umanist Staffing • Pune District

On-site
INR 2,250,000 - 2,750,000
DevOps Engineer/Site Reliability Engineer
DevOps Engineer/Site Reliability Engineer

Thompsons HR Consulting Pvt Ltd • Pune District

On-site
INR 900,000 - 1,300,000
Software Engineer
Software Engineer

PwC • Hyderabad, Bengaluru

Hybrid
INR 2,800,000 - 5,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Skillventory • Kamrup Metropolitan

On-site
INR 1,400,000 - 2,200,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AcquireX • Pune District

On-site
INR 1,200,000 - 1,800,000
Health insurance
Flexible working hours
Training opportunities
Senior Site Reliability Engineer
Senior Site Reliability Engineer

VMC Soft Technologies, Inc • Hyderabad

On-site
INR 1,500,000 - 2,000,000