Manager Technology

wk

India

On-site

INR 4,000,000 - 7,000,000

Full time

8 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

wk is seeking a Manager, Technology with 12+ years of experience in SRE, platform engineering, and infrastructure. You will lead a 6-8 person team across SRE and platform operations, drive reliable scalable platform solutions, and partner with architects and product teams to define roadmaps and implement best practices.

You will shape platform reliability through SRE practices, monitoring, SLOs, and automation, while leveraging AI tools for faster IaC and incident response.

Qualifications

  • Bachelor's degree in information technology or related field.
  • Preferred 12+ years' experience in SRE, Platform Engineering, or Infrastructure.
  • Extensive Agile experience and involvement in roadmaps.
  • Hands-on Kubernetes (AKS preferred) administration at scale.
  • Proficient with Terraform, ArgoCD & Argo Workflows, Helm/Kustomize.
  • Experience with GitOps and CI/CD pipelines (GitHub Actions, Azure DevOps).
  • Driving adoption of platform capabilities and onboarding engineering teams.
  • Defining and managing SLOs, SLIs, and error budgets.
  • Production operations at scale: capacity planning, DR, failover.
  • Building internal developer platforms and self-service tooling.
  • Toil reduction through automation and self-healing systems.
  • Familiar with observability tools: Datadog, Prometheus, Grafana.

Responsibilities

  • Experience in platform release cycles, SRE best practices, infra code reviews, and incident/defect management.
  • Lead a team of 6-8 engineers across SRE, Platform Operations, and platform capabilities functions.
  • Design scalable platform solutions addressing reliability, operability, and developer experience.
  • Establish SRE practices including monitoring, alerting, SLOs, and performance tuning across production infrastructure.
  • Oversee day-2 operations for AKS clusters, including upgrades and capacity planning.
  • Identify optimal technologies to improve platform reliability and developer self-service.
  • Create POCs and collaborate with architects to build technical roadmaps.
  • Drive Spec Driven Development (SDD) for platform services and infra contracts.
  • Leverage AI tools to accelerate infrastructure-as-code and incident response.
  • Maintain platform documentation, runbooks, and operational playbooks.
  • Mentor team members; provide training, advice, coaching, and opportunities.
  • Promote AI tooling usage in SDLC and SRE workflows.

Skills

SRE
Platform Engineering
Team Leadership
Agile Experience
Infra Automation

Education

Bachelor's degree in information technology or related field

Tools

Kubernetes (AKS)
Terraform
ArgoCD
Argo Workflows
Helm/Kustomize
GitOps
GitHub Actions
Azure DevOps
Datadog
Prometheus
Grafana
CI/CD Pipelines

Job description

Role:Manager, Technology. Experience Required: 12+ years
Responsibilities
  • Experience in platform release cycles, SRE best practices, infrastructure code reviews, and incident/defect management.
  • Efficient in handling a team of 6-8 engineers across SRE, Platform Operations, and platform capabilities functions.
  • Out of the box thinking and creative problem-solving skills is desired.
  • Works with architects and product owners/managers to design and implement scalable platform solutions addressing reliability, operability, and developer experience.
  • Works with engineering teams to establish SRE practices including monitoring, alerting, SLOs, and performance tuning across production infrastructure.
  • Manages day-2 operations for AKS clusters, including twice-yearly upgrades, patching, and capacity planning.
  • Identifies optimal technologies and practices to improve platform reliability and developer self-service. Stays current with the evolving cloud-native and platform engineering space to evaluate tools and capabilities for integration.
  • Involved in creating POCs, interacting with architects within groups to strategize platform development and build technical roadmaps.
  • Drives Spec Driven Development (SDD) practices for platform capabilities and SRE automation - defining clear specifications for infrastructure contracts, APIs, and platform services before implementation to ensure predictable, well-documented outcomes.
  • Leverages AI tools (Co-pilot, Claude, agentic workflows) to accelerate infrastructure-as-code development, incident triage, runbook automation, and platform self-service capabilities.
  • Helps in maintaining proper platform documentation, runbooks, and operational playbooks.
  • Supports and mentors team members by providing training, advice, coaching, and educational opportunities.
  • Promotes culture of using various AI tools in SDLC including Co-pilot, Claude, etc.
Job Qualifications
  • Bachelor's degree in information technology or related field.
  • Preferred 12+ years' experience in SRE, Platform Engineering, or Infrastructure domain.
  • Extensive experience working in Agile, participating in various L0, technical design, and roadmap initiatives.
  • Strong hands‑on experience required in Kubernetes (AKS preferred). Should be proficient in operating and managing Kubernetes clusters at an organization level.
  • Strong experience in Terraform, ArgoCD & Argo Workflows, Helm/Kustomize, and GitOps practices.
  • Experience with CI/CD pipelines (GitHub Actions, Azure DevOps) and infrastructure-as-code at scale.
  • Strong experience in driving adoption of platform capabilities for developers, setting up COPs, and driving onboarding of engineering teams onto the platform.
  • Proven experience defining and managing SLOs, SLIs, and error budgets to balance reliability with feature velocity across production systems.
  • Experience building and running incident management processes including on‑call rotations, escalation frameworks, blameless post‑mortems, and RCAdet (???)
  • Hands‑on experience with production operations at scale - capacity planning, disaster recovery, failover strategies, and business continuity for cloud-hosted infrastructure.
  • Experience building internal developer platform capabilities - self-service tooling, golden paths, developer portals, and "platform as a product" approaches that reduce toil for engineering teams.
  • Demonstrated ability to drive toil reduction through automation - scripting operational tasks, building self-healing systems, and implementing proactive alerting to reduce reactive firefighting.
  • Experience conducting Production Readiness Reviews (PRRs) and defining operational standards for services transitioning to production.
  • Experience with Spec Driven Development (SDD) methodologies - defining infrastructure and platform service specifications upfront to drive implementation, testing, and validation of platform capabilities.
  • Familiarity with AI-assisted development tools and agentic automation patterns applied to SRE workflows such as intelligent alerting, automated remediation, and infrastructure provisioning.
  • Passionate about sharing your experiences and knowledge and growing your team.
  • Ability to creatively handle challenges and obstacles, innovating solutions balancing both immediate needs with longer-term ownership and maintenance.
  • Preferred experience with observability tooling (Datadog, Prometheus, Grafana) for platform monitoring.
  • Preferred knowledge in microservices architecture, service mesh, and cloud-native patterns.
  • Strong interpersonal and communication skills, coupled with solid teamwork ethic and customer focus.
Our Interview Practices

To maintain a fair and genuine hiring process, we kindly ask that all candidates participate in interviews without the assistance of AI tools or external prompts. Our interview process is designed to assess your individual skills, experiences, and communication style. We value authenticity and want to ensure we're getting to know you-not a digital assistant. T

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Nexcess • India

On-site
INR 2,000,000 - 4,200,000
Lead DevOps / Platform Engineering Lead
Lead DevOps / Platform Engineering Lead

Wolters Kluwer • Pune District

On-site
INR 3,500,000 - 7,000,000
SRE Architect
SRE Architect

Prodapt • Chennai District

On-site
INR 6,000,000 - 9,000,000
Senior Manager - Site Reliability Engineer|NR-2026-0246
Senior Manager - Site Reliability Engineer|NR-2026-0246

Media.net • Bengaluru

On-site
INR 6,000,000 - 8,000,000
Platform Architect
Platform Architect

r3 Consultant • Dadri

On-site
INR 2,400,000 - 4,800,000
Platform SRE
Platform SRE

YASH Technologies • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Director Cloud & Infrastructure Architect (Multi-Cloud | Datacenter | SRE)
Director Cloud & Infrastructure Architect (Multi-Cloud | Datacenter | SRE)

Mancer Consulting Services • Bengaluru

On-site
INR 3,500,000 - 6,500,000
Lead SDE - DevOps
Lead SDE - DevOps

Flourish Ventures • Chennai District

On-site
INR 2,000,000 - 3,000,000
Inclusive and people-first culture
Health & wellness programs
Comprehensive medical insurance
+2
Senior Director of DevOps
Senior Director of DevOps

GreyOrange • Gurugram District

On-site
INR 6,000,000 - 9,000,000
Senior Platform Engineer – AWS / Kubernetes / DevOps
Senior Platform Engineer – AWS / Kubernetes / DevOps

Inadev India • Kolkata District

On-site
INR 4,000,000 - 6,000,000