Senior Kubernetes Engineer for AI Infrastructure (Remote)

HCLTech

United States

On-site

USD 78,000 - 148,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

HCLTech is seeking a Senior Kubernetes Administrator – AI Infrastructure to lead platform engineering for AI workloads in a remote setting. You will build, troubleshoot, and scale Kubernetes platforms supporting model development, training, and inference services.

Ideal candidates have 7+ years in infrastructure, deep Kubernetes expertise, experience with GPUs, Helm, GitOps, and strong Linux networking. Remote work with competitive compensation and benefits.

Qualifications

  • 7+ years in infrastructure engineering with hands-on Kubernetes administration.
  • Deep understanding of Kubernetes internals and cluster troubleshooting.
  • Experience with container runtimes, Helm, GitOps, and declarative operations.
  • Experience supporting GPU workloads on Kubernetes in lab/production.
  • Strong Linux administration and data center networking knowledge.
  • Ability to debug from symptom to root cause across node, pod, network, storage and control plane.
  • Scripting in Python, Bash, or Go.

Responsibilities

  • Build, administer, and troubleshoot Kubernetes platforms used for AI and data-intensive workloads.
  • Diagnose failures across control plane components, kubelet, CNI, CSI, ingress, and scheduling.
  • Support GPU-enabled Kubernetes environments, including device plugin behavior and drivers.
  • Improve platform reliability through automation, configuration standards, upgrade planning, and validation gates.
  • Investigate storage throughput, network policy, DNS, image pulls, autoscaling, and degraded node states.
  • Collaborate with Linux, network, validation, and SRE teams to resolve cross-layer issues.
  • Create reusable runbooks, dashboards, and health checks for day-2 operations.
  • Contribute to platform hardening and tenant readiness.

Skills

Kubernetes admin
Linux admin
Python
Bash
Go
GPU workloads
Cluster troubleshooting
GitOps
Helm
Declarative ops
Networking basics

Tools

Kubeflow
Argo
Prometheus
Grafana
Loki
Service mesh
Bare-metal Kubernetes
High-performance storage

Job description

HCLTech is seeking a Senior Kubernetes Administrator – AI Infrastructure to lead platform engineering for AI workloads in a remote setting. You will build, troubleshoot, and scale Kubernetes platforms supporting model development, training, and inference services.

Ideal candidates have 7+ years in infrastructure, deep Kubernetes expertise, experience with GPUs, Helm, GitOps, and strong Linux networking. Remote work with competitive compensation and benefits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Kubernetes Administrator – AI Infrastructure
Senior Kubernetes Administrator – AI Infrastructure

HCLTech • United States

On-site
USD 78,000 - 148,000
Kubernetes Administrator – AI Infrastructure
Kubernetes Administrator – AI Infrastructure

Sira Consulting, an Inc 5000 company • United States

On-site
USD 140,000 - 210,000
AI Infrastructure Engineer - GPU & Kubernetes
AI Infrastructure Engineer - GPU & Kubernetes

HCLTech • California (MO)

On-site
USD 150,000 - 210,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
Remote Senior Kubernetes Platform Engineer - AI & GPUs
Remote Senior Kubernetes Platform Engineer - AI & GPUs

Sira Consulting, an Inc 5000 company • United States

On-site
USD 140,000 - 210,000
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote
Senior AI Infrastructure Engineer — Scale GPU Clusters Remote

AI Breaking Wire • San Francisco (CA), Northern (KY)

Hybrid
USD 280,000 - 420,000
Equity options
Health, vision, dental benefits
Unlimited PTO
+2
Senior Kubernetes Cloud Engineer (AI Platform)
Senior Kubernetes Cloud Engineer (AI Platform)

Hewlett Packard Enterprise • Houston (TX)

On-site
USD 93,000 - 214,000
Health & Wellbeing
Personal & Professional Development
Unconditional Inclusion
Senior Kubernetes & GPU Infra Engineer for AI-scale Compute
Senior Kubernetes & GPU Infra Engineer for AI-scale Compute

Kindredventures • United States

On-site
USD 140,000 - 190,000
Senior Cloud Developer – Kubernetes AI Platform
Senior Cloud Developer – Kubernetes AI Platform

Hewlett Packard Enterprise • Durham (NC)

On-site
USD 93,000 - 214,000
Senior AI Infra Engineer – Cloud, Kubernetes & OSS (Remote)
Senior AI Infra Engineer – Cloud, Kubernetes & OSS (Remote)

Soloioinc • United States

On-site
USD 100,000 - 130,000
AI Infrastructure Engineer — GPU Kubernetes for Production
AI Infrastructure Engineer — GPU Kubernetes for Production

vCluster • Germany (OH)

On-site
USD 150,000 - 200,000
Competitive Salary
Platinum-Level Insurance
Flexible Working Schedule
+1