Production Engineering Manager

CoreWeave

Sunnyvale (CA)

On-site

USD 180,000 - 250,000

Full time

11 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

CoreWeave is seeking an Engineering Manager to build and lead a team of infrastructure and systems engineers. You will shape the team at its earliest stage, translating an evolving charter into a clear, sustainable operating model.

In partnership with tech leadership, you will define how the team delivers, operates, and grows while owning its people, priorities, and outcomes. You will lead with focus on reliability, observability, testing, and safe change practices, while driving architecture

Qualifications

  • 6+ years of experience in software, infrastructure, platform, SRE, or production engineering with people leadership.

Responsibilities

  • Lead and develop a high-performing team of infrastructure engineers.
  • Set clear priorities, roadmaps, ownership, and measures of success.
  • Define decision-making, escalation, and on-call practices as the team grows.
  • Partner with technical leaders to guide architecture and deliver reliable infrastructure and automation.
  • Establish strong practices for testing, observability, incident response, and safe changes.

Skills

Leadership
Roadmap planning
Communication
Cloud infrastructure
Observability
SRE practices
On-call management
Kubernetes
Production systems

Tools

Kubernetes
GitOps
Observability tools

Job description

  • Prime Systems & Services builds and operates internal Kubernetes platforms and foundational infrastructure services that CoreWeave depends on. We are a new team focused on making these systems reliable, scalable, and easy to operate. Our work combines Kubernetes, software, systems, and reliability engineering to improve how infrastructure is deployed, changed, and supported at scale
  • We are seeking an Engineering Manager to build and lead a team of infrastructure and systems engineers. You will have the opportunity to shape the team from its earliest stage and translate an evolving charter into a clear, sustainable operating model
  • In partnership with technical leadership, you will establish how the team delivers, operates, and grows while remaining accountable for its people, priorities, and outcomes
  • Lead and develop a high-performing team of infrastructure engineers
  • Set clear priorities, roadmaps, ownership, and measures of success
  • Define clear decision-making, escalation, and on-call practices as the team grows
  • Partner with technical leaders to guide architecture and deliver reliable infrastructure and automation
  • Establish strong practices for testing, observability, safe changes, incident response, and on-call
  • Use operational and delivery data to identify risks and guide investment
  • Resolve cross-team dependencies and create a culture of accountability, collaboration, and continuous improvement

You have experience hiring, coaching, managing performance, and communicating priorities and risks to technical teams and senior leadersYou have experience with Kubernetes, distributed systems, and cloud infrastructureYou have managed engineers responsible for production systems with meaningful reliability requirementsYou can guide technical tradeoffs and work effectively with senior engineers without becoming the team’s sole architectYou have 6+ years of experience in software, infrastructure, platform, SRE, or production engineering, including people leadershipYou have turned ambiguous ownership or cross-team dependencies into clear plans and delivered outcomesYou have improved operational practices such as on-call, SLOs, observability, incident response, or change managementWe value diverse experiences and do not expect candidates to match every qualification. If you enjoy developing engineers, creating clarity, and building dependable infrastructure organizations, we would love to talkFamiliarity with infrastructure automation, GitOps, and progressive deliveryExperience building a new engineering team or platform functionExperience leading teams with both development and operational responsibilitiesExperience operating Kubernetes at scale or on bare metal

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineering Manager, Production Infrastructure & Kubernetes
Engineering Manager, Production Infrastructure & Kubernetes

CoreWeave • Bellevue (WA)

On-site
USD 182,000 - 242,000
Medical insurance
401(k) with employer match
Stock options
Engineering Manager, Kubernetes & Platform Reliability
Engineering Manager, Kubernetes & Platform Reliability

CoreWeave Europe • Sunnyvale (CA)

On-site
USD 182,000 - 242,000
Healthcare
ESPP
401(k) match
+9
Kubernetes Infrastructure Engineering Manager
Kubernetes Infrastructure Engineering Manager

CoreWeave • Sunnyvale (CA)

On-site
USD 180,000 - 250,000
Production Infrastructure Lead (Kubernetes & Scale)
Production Infrastructure Lead (Kubernetes & Scale)

Socket.dev • New York (NY), Sunnyvale (CA), Bellevue (WA)

On-site
USD 182,000 - 242,000
Medical insurance
401(k) match
Flexible PTO
+5
IT Systems Engineer: Cloud Infrastructure
IT Systems Engineer: Cloud Infrastructure

CoreWeave • Sunnyvale (CA)

On-site
USD 150,000 - 190,000
IT Systems Engineer: Cloud Infrastructure
IT Systems Engineer: Cloud Infrastructure

CoreWeave • Bellevue (WA)

On-site
USD 150,000 - 210,000
Staff Platform Engineer: Kubernetes & Data Infra
Staff Platform Engineer: Kubernetes & Data Infra

CoreWeave • Livingston (NJ)

On-site
USD 180,000 - 280,000
Medical insurance
401(k) match
Paid time off
+2
Staff Platform Engineer, Kubernetes & Data Infra
Staff Platform Engineer, Kubernetes & Data Infra

CoreWeave • New York (NY)

On-site
USD 190,000 - 260,000
Healthcare: medical, dental, vision
Life Insurance (company-paid)
Disability insurance
+6
Platform Engineer
Platform Engineer

Synergy • Chicago (IL)

On-site
USD 100,000 - 150,000
Staff Software Engineer, Kubernetes Platform & Reliability
Staff Software Engineer, Kubernetes Platform & Reliability

CoreWeave • Bellevue (WA)

On-site
USD 180,000 - 240,000
Medical Insurance
Dental Insurance
Vision Insurance
+6