Staff Infra Engineer: AI-Scale Cloud & Kubernetes

LlamaIndex, Inc.

San Francisco (CA)

Hybrid

USD 200,000 - 275,000

Full time

3 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive base salary and equity
Comprehensive medical/dental/vision
Unlimited paid time off
Daily catered lunch and snacks

Job summary

LlamaIndex, Inc. in San Francisco is seeking an experienced Infra engineer to design, build, and scale core infrastructure powering a high-volume data platform for AI applications.

You will manage cloud resources, Kubernetes clusters, and deployment tooling, collaborating with multiple engineering teams to support rapid growth and diverse deployment models. This role emphasizes cost efficiency, security, and observability.

Qualifications

  • 8+ years of engineering experience.
  • Experience on Platform or Infrastructure teams with major projects involving infrastructure components (Terraform/CDKTF, Kubernetes, Helm, test infrastructure, release management, observability).
  • Experience in optimizing cloud resource utilization.
  • Proficient in tuning Kubernetes clusters and cloud resources for cost and performance efficiency.
  • Willing to build LlamaIndex’s engineering culture as we grow.
  • You can balance speed and pragmatism and build the appropriate solutions for each stage of the company’s growth.
  • Hands-on proficiency with modern LLM tooling and production AI systems, including model APIs, agent or RAG frameworks, evaluation and tracing tools, and the operational characteristics of LLM workloads.

Responsibilities

  • Collaborate with other engineering teams to build and maintain foundational systems that empower developers and support rapid growth.
  • Design and implement scalable infrastructure solutions for deployment models including SaaS, single-tenant, and private deployments.
  • Manage and optimize cloud resources and Kubernetes clusters for cost-effectiveness and performance.
  • Enable external customer deployment success through maintaining clear infrastructure boundaries and principles.
  • Optimize and improve the release and deployment processes to enhance efficiency and reliability.
  • Ensure compliance with relevant regulations and implement robust security measures across environments.
  • Build and operate infrastructure for production LLM applications, including model integrations, inference workloads, evaluation pipelines, observability, and agentic workflows.

Skills

Platform/infrastructure experience
Solving complex systems
Balancing speed and pragmatism
LLM tooling familiarity

Tools

Terraform/CDKTF
Kubernetes
Helm
Observability
Release management

Job description

LlamaIndex, Inc. in San Francisco is seeking an experienced Infra engineer to design, build, and scale core infrastructure powering a high-volume data platform for AI applications.

You will manage cloud resources, Kubernetes clusters, and deployment tooling, collaborating with multiple engineering teams to support rapid growth and diverse deployment models. This role emphasizes cost efficiency, security, and observability.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Infra Platform Engineer — Scale Cloud & Kubernetes
Infra Platform Engineer — Scale Cloud & Kubernetes

LlamaIndex • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Competitive base salary and equity compensation
Comprehensive benefits coverage
Unlimited paid time off
+1
Staff Infra Engineer — Scale Global Kubernetes & AI Infra
Staff Infra Engineer — Scale Global Kubernetes & AI Infra

Icehouseventures • San Francisco (CA)

On-site
USD 287,000 - 485,000
Competitive pay
Equity
Relocation assistance
+7
Scale AI Infra Engineer | Kubernetes & CI/CD
Scale AI Infra Engineer | Kubernetes & CI/CD

BaseTen • New York (NY), San Francisco (CA)

Hybrid
USD 160,000 - 210,000
Meaningful equity
100% medical, dental, vision coverage
Flexible PTO including Winter Break
+4
AI Infra Platform Engineer: Scale Kubernetes & AI Workloads
AI Infra Platform Engineer: Scale Kubernetes & AI Workloads

BrainCo • San Francisco (CA)

On-site
USD 100,000 - 140,000
Competitive salary plus equity
Daily lunches
Commuter benefits
+3
Senior Platform Engineer, AI Scale & Cloud Foundations
Senior Platform Engineer, AI Scale & Cloud Foundations

Scale AI, Inc. • San Francisco (CA)

On-site
USD 216,000 - 270,000
Equity
Health, dental & vision
Retirement benefits
+3
Cloud Infrastructure Engineer — Kubernetes & Scale
Cloud Infrastructure Engineer — Kubernetes & Scale

OpenAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Relocation assistance
Senior DevOps Engineer - AI Infra, Kubernetes & Cloud
Senior DevOps Engineer - AI Infra, Kubernetes & Cloud

StackAI • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Hybrid work model
Remote work available
Office near Salesforce Park
AI Infra Engineer: Scale ML Clusters on Kubernetes & Slurm
AI Infra Engineer: Scale ML Clusters on Kubernetes & Slurm

Perplexity • Palo Alto (CA)

On-site
USD 220,000 - 405,000
Senior Infra Engineer: Scale Global Kubernetes & LLM Infra
Senior Infra Engineer: Scale Global Kubernetes & LLM Infra

Icehouseventures • San Francisco (CA)

On-site
USD 249,000 - 405,000
Competitive Compensation
Equity
Relocation and Visa Support
+6
AI Infra Engineer: Kubernetes & LLM Ops
AI Infra Engineer: Kubernetes & LLM Ops

Alignity Solutions • New York (NY)

On-site
USD 120,000 - 180,000