Sr. Site Reliability Engineer

Visa Consolidated Support Services India

Bengaluru

On-site

INR 1,500,000 - 2,500,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Visa Consolidated Support Services India is seeking an engineer responsible for leading infrastructure orchestration and automation. You will own the lifecycle of core platform components, ensuring their reliability and operational excellence through SRE practices. Required qualifications include expertise in public cloud platforms, Kubernetes at scale, and Infrastructure as Code. Strong collaboration skills are essential for this role. Join us in Bengaluru to drive innovative solutions in cloud infrastructure.

Qualifications

  • Experience administrating productive Kubernetes environments.
  • Proven understanding of Service Mesh technologies.
  • Familiarity with observability tooling and incident management.

Responsibilities

  • Own the end-to-end lifecycle of core platform components.
  • Lead the design and implementation of infrastructure bootstrap orchestration.
  • Apply SRE practices to enhance platform reliability and operational excellence.

Skills

Public Cloud platforms (AWS preferred, Azure)
Kubernetes at scale
Service Mesh technologies (e.g., Istio preferred)
Observability tooling and Golden Signals concepts
Incident management concepts
Infrastructure as Code (e.g., Terraform)
Cloud-Native containerized micro-services architecture
Strong collaboration and communication skills

Job description

This engineer is expected to lead by example through hands-on contributions, deep technical expertise, and cross-team influence, particularly in the area of infrastructure bootstrap orchestration and automation at scale.

Key Responsibilities
Platform Ownership & Reliability

Own the end-to-end lifecycle (design, provisioning, upgrades, and decommissioning) of core platform components, including:

  • Cloud infrastructure primitives
  • Kubernetes clusters and cluster services
  • Networking, ingress, and service discovery
  • Service Mesh and supporting data-plane components

Ensure platform components are resilient by design, applying SRE principles such as:

  • Fault isolation and graceful degradation
  • Capacity planning and saturation control
  • Reduced operational toil and clear failure modes
  • Continuously assess and mitigate reliability risks, proactively improving platform stability and operational readiness.
Infrastructure Bootstrap & Automation Leadership

Lead the design and implementation of infrastructure bootstrap orchestration, including:

  • Automated cluster and environment provisioning
  • Deterministic, repeatable platform bring-up and teardown
  • Dependency-aware orchestration across cloud, network, and Kubernetes layers

Drive a strong Infrastructure-as-Code and GitOps-first approach, ensuring:

  • Platform components are reproducible and auditable
  • Changes are automated, testable, and reversible
  • Manual intervention is minimized or eliminated
  • Identify automation gaps and lead initiatives that significantly reduce human effort, onboarding time, and operational risk.
SRE Practices & Operational Excellence

Apply and promote SRE practices across the platform, including:

  • Clear ownership and runbooks for platform components
  • Participation in on-call rotation as a platform reliability escalation point
  • Incident response, post-incident reviews, and problem management

Improve platform operability by:

  • Simplifying day-2 operations
  • Standardizing upgrade and rollback strategies
  • Reducing Mean Time to Detect (MTTD) and Mean Time to Recover (MTTR)
  • Ensure platform operations align with security, compliance, and internal control requirements.
Qualifications
  • Public Cloud platforms (AWS preferred, Azure)
  • Kubernetes at scale, previous experience administrating productive Kubernetes environments
  • Service Mesh technologies (e.g., Istio preferred, App Mesh, Linkerd)
  • Observability tooling and Golden Signals concepts
  • Incident management concepts and on-call operations
  • Infrastructure as Code (e.g., Terraform)
  • Cloud-Native containerized micro-services architecture
  • Strong collaboration and communication skills
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Cloud Site Reliability Engineer
Senior Cloud Site Reliability Engineer

Augusta Infotech • Bengaluru

Hybrid
INR 1,500,000 - 2,500,000
Site Reliability Engineer
Site Reliability Engineer

Smart Ims • Bengaluru

Hybrid
INR 1,200,000 - 2,000,000
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Sierra Ventures • Bengaluru

On-site
INR 3,500,000 - 5,500,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Headout • Bengaluru

On-site
INR 1,200,000 - 1,800,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Zorba AI • Chennai District

On-site
INR 1,200,000 - 2,400,000
Technology Manager
Technology Manager

Wolters Kluwer • Pune District

Hybrid
INR 2,800,000 - 5,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Falabella India • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Five9 • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Senior SRE
Senior SRE

CloudRaft • India

On-site
INR 2,500,000 - 4,500,000
Competitive salary
Premium health insurance & wellness
AI stack & GPU infrastructure
+2
SRE Engineer
SRE Engineer

Prodapt Solutions Private Limited • Chennai District

On-site
INR 1,800,000 - 3,000,000