Devops Lead

Artech Infosystems Private Limited

India

Remote

INR 3,500,000 - 7,500,000

Full time

3 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Artech Infosystems Private Limited is seeking a Senior DevOps Engineer to join our Infrastructure team and lead Kubernetes operations at scale. You will own health, reliability, and automation across multi-cluster environments, designing tooling that platform teams rely on to manage infrastructure safely.

You will mentor junior engineers, drive incident response, and collaborate with service owners to improve observability, CI/CD pipelines, and deployment strategies across AWS/GCP/Azure.

Qualifications

  • 10+ years of DevOps, SRE, or Infrastructure Engineering experience.
  • Hands-on Kubernetes operations across large multi-cluster environments.
  • Strong programming skills in Python, Go, or Bash for automation.
  • Experience with cloud providers (AWS/GCP/Azure) and cost/performance tradeoffs.
  • Excellent written and verbal communication; mentoring experience.

Responsibilities

  • Own roadmap for Kubernetes cluster operations across environments.
  • Lead troubleshooting of complex production and non-production issues.
  • Architect and lead internal tooling for automation and operator development.
  • Drive automation strategies for upgrade, remediation, and scaling workflows.
  • Diagnose infrastructure blockers and improve observability with logs/metrics.
  • Set standards for issue tracking, triage, and post-incident reviews.
  • Collaborate with service owners, platform teams, and leadership.
  • Lead on-call rotations and mentor junior engineers.

Skills

Kubernetes operations
SRE experience
Python/Go/Bash scripting
Cloud infrastructure (AWS/GCP/Azure)
Debugging distributed systems
Strong communication
Mentoring/leadership

Tools

Terraform
Helm
Ansible
Jenkins
GitHub Actions
Spinnaker
ArgoCD
Flux
Datadog
Prometheus

Job description

Remote

Immediate Joiners Only

Shift Timing: 07:30 PM To 04:30 AM

ABOUT THE ROLE

We are looking for a Senior DevOps Engineer to join our Infrastructure team, focused on Kubernetes operations and automation at scale. You will drive the strategy and own the health, scalability, and reliability of our Kubernetes clusters,nwhile architecting and leading the development of tooling that platform and service teams rely on to manage infrastructure safely and efficiently. You will work closely with service owners, platform engineers, and tooling teams — and mentor junior engineers — to keep clusters running smoothly and to automate away manual, error-prone operational work.

WHAT YOU'LL DO
  • Own and drive the roadmap for Kubernetes cluster operations — lead cluster upgrades, manage node groups (scaling, draining, replacement), and maintain overall cluster health across environments at scale
  • Lead troubleshooting of complex production and non-production issues — use kubectl, logs, metrics, and other diagnostics to identify root cause and resolve workload, node, and networking failures, often across multiple interdependent systems
  • Architect and lead development of internal tooling — design, implement, and maintain automation ( s, CLIs, controllers, operators) that helps teams manage infrastructure more reliably and with less manual effort
  • Drive automation strategy for operational workflows — replace manual runbooks with s and tools that handle upgrades, remediation, scaling, and routine maintenance
  • Diagnose and resolve complex infrastructure blockers — debug deployment failures, node/pod scheduling issues, resource constraints, and misconfigurations across clusters
  • Define and improve observability strategy — instrument logging, metrics, and alerting to increase visibility into cluster health and reduce time-to-detect/resolve
  • Set standards for issue tracking and triage — author detailed bug reports capturing root cause, repro steps, and impact; drive issues to resolution and improve team-wide reporting practices
  • Partner cross-functionally with service owners, platform teams, and leadership — align on operational requirements, capacity planning, and upgrade/maintenance schedules
  • Lead on-call rotations and incident response, own post-incident reviews, and drive continuous improvement initiatives across the infrastructure org
  • Mentor and upskill junior and mid-level engineers, providing technical guidance and code/design reviews
WHAT WE'RE LOOKING FOR
Required
  • 10+ years of DevOps, SRE, or Infrastructure Engineering experience
  • Deep, hands-on Kubernetes operations expertise — cluster upgrades, node group management, and complex troubleshooting across large-scale, multi-cluster environments
  • Strong programming/ ing skills (Python, Go, Bash, or similar) — proven track record building, scaling, and maintaining automation and internal tooling used by multiple teams
  • Advanced skills in reading and interpreting logs, metrics, and system state to diagnose complex, cross-system infrastructure issues
  • Extensive experience with cloud infrastructure (AWS, GCP, or Azure), including architecture decisions and cost/performance tradeoffs
  • Expert-level debugging skills across distributed systems — able to trace failures from symptom through to root cause in highly complex environments
  • Excellent written and verbal communication — able to produce clear status updates, bug reports, technical documentation, and influence technical direction across teams
  • Demonstrated experience mentoring engineers and leading technical initiatives
Preferred
  • Deep experience with infrastructure-as-code tools (Terraform, Helm, Ansible), including designing reusable modules/patterns for org-wide use
  • Strong familiarity with CI/CD systems (Jenkins, GitHub Actions, Spinnaker, or similar), including pipeline architecture Hands-on experience with GitOps workflows (ArgoCD, Flux) at scale
  • Advanced experience with observability stacks (Prometheus, Grafana, Datadog), including designing alerting/SLO frameworks
  • Proven experience operating Kubernetes at scale across multiple clusters, regions, or environments, including capacity planning and disaster recovery
  • Experience contributing to or leading architectural decisions for infrastructure platforms
TECHNOLOGIES YOU'LL WORK WITH

Kubernetes kubectl Python Terraform Helm Prometheus/Grafana GitHub CI/CD tooling Cloud infrastructure (AWS/GCP/Azure).

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Devops Engineer
Senior Devops Engineer

Artech L.L.C. • India

On-site
INR 4,000,000 - 7,000,000
DevOps Engineer
DevOps Engineer

COZZERA INTERNATIONAL LLP • Dadri

On-site
INR 1,600,000 - 2,000,000
Sr. Devops Lead
Sr. Devops Lead

Artech Infosystems Private Limited • India

Remote
INR 4,000,000 - 7,000,000
Senior DevOps Engineer
Senior DevOps Engineer

Coredge • Dadri

On-site
INR 1,000,000 - 1,500,000
DevOps Engineer
DevOps Engineer

UsefulBI Corporation • Bengaluru

On-site
INR 1,800,000 - 3,200,000
DevOps Engineer
DevOps Engineer

Unicommerce • Gurugram District

On-site
INR 1,200,000 - 1,800,000
Senior DevOps Engineer
Senior DevOps Engineer

Broadway Global Services • Bengaluru

On-site
INR 1,800,000 - 3,200,000
Devops Engineer
Devops Engineer

Euphoric Thought Technologies Pvt. Ltd. • Bengaluru

On-site
INR 2,500,000 - 3,800,000
Senior Manager - DevOps
Senior Manager - DevOps

Greythr • Bengaluru

On-site
INR 4,000,000 - 6,000,000
Senior DevOps Engineer
Senior DevOps Engineer

BreachLock Inc. • Pune District

On-site
INR 1,500,000 - 2,000,000
Competitive compensation with performance-linked incentives
Flexible working arrangements
Direct visibility with leadership