Senior Specialist Field Engineer - Compute Infrastructure

CoreWeave

Sunnyvale (CA)

On-site

USD 130,000 - 170,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

CoreWeave is seeking a Specialist Field Engineer - Compute Infrastructure in Sunnyvale, California. The role focuses on optimizing GPU cluster functionalities for major customers, transforming raw hardware into reliable compute solutions. In this position, you will lead technical projects and collaborate with multiple departments to enhance CoreWeave’s offerings and build strong client relationships.

The ideal candidate will possess over 7 years in a relevant field, have hands-on experience with GPU clusters, and understand cloud infrastructure deeply. This is an opportunity to contribute to cutting-edge AI technologies while working directly with clients to ensure successful deployments.

Qualifications

  • 7+ years in Solutions Architect or similar roles in Cloud Infrastructure.
  • Hands-on experience with large GPU clusters and bare-metal provisioning.
  • Expertise in modern rack-scale GPU server hardware.

Responsibilities

  • Own technical path from design to production-ready supercomputer.
  • Lead GPU cluster validation and performance benchmarking.
  • Establish strong relationships with customers for success.

Skills

GPU cluster management
Cloud Infrastructure
Linux system administration
Networking fundamentals
Customer relationship management

Education

B.S. in Computer Science or related field

Tools

Kubernetes
Ansible
Python

Job description

  • The Field Engineering organization at CoreWeave is dedicated to ensuring every customer running AI workloads at scale has a seamless, reliable, and high-performance experience
  • This team supports the infrastructure that powers the AI revolution—working across data centers, hardware systems, and customer workloads to maintain the integrity of our cloud platform
  • Field Engineering aligns closely with internal and customer engineering teams, offering valuable insights from the field and the chance to shape the CoreWeave product roadmap and development
  • As a Specialist Field Engineer - Compute Infrastructure at CoreWeave, you’ll own the technical path for some of our largest customers as they go from facility and rack design to a validated, production-ready supercomputer
  • Working alongside the teams that build and operate each layer, you are the deep technical expert who turns raw data center hardware—racks, GPUs, high-speed fabric, firmware—into reliable compute that customers can train and inference on at scale, spanning infrastructure engineering, provisioning, validation, operations, and support
  • You’ll engage hands‑on across the entire customer lifecycle: leading new GPU cluster bring‑up and acceptance, driving InfiniBand/RoCE fabric validation and HPC performance benchmarking, defining how we operate customer bare‑metal fleets at rack‑level-and-up (IT service, break‑fix, network, and firmware), and standing up locked‑down, security‑sensitive environments for our most strategic AI customers
  • You’ll partner closely with Data Center Operations, Fleet Operations, Networking, and Product Engineering, and your work in the field will directly shape how CoreWeave delivers compute infrastructure
  • Serve as the primary technical point of contact for customers, establishing strong technical relationships and ensuring their success with CoreWeave’s cloud infrastructure offerings, focusing on bare‑metal compute infrastructure and end‑to‑end cluster delivery within high‑performance compute (HPC) environments
  • Own the technical path from facility and rack design to a validated, production‑ready supercomputer—spanning logical design, infrastructure engineering, provisioning, validation, operations, and support
  • Lead bring‑up and acceptance of new large‑scale GPU clusters, driving InfiniBand/RoCE fabric validation, HPC performance benchmarking (e.g., NCCL, ib_write_bw), and remediation of fabric, optics, firmware, and node‑level issues to meet customer performance targets
  • Define and operationalize models for managing customer bare‑metal fleets at rack‑level-and-up—IT service, break‑fix, network and firmware management—including Bare Metal as a Service (BMaaS) and customer self‑service patterns
  • Partner with Data Center Operations, Fleet Operations, and Networking teams to align facility, hardware, and fabric readiness with customer go‑live timelines and operational SLAs
  • Review and advise on customer‑facing technical contract terms, including service scope, operational responsibilities, SLAs, isolation requirements, and support boundaries
  • Lead proof of concept initiatives to showcase the value and viability of CoreWeave’s solutions within specific environments
  • Drive technical leadership and direction during customer meetings, presentations, and workshops, addressing any technical queries or concerns that arise
  • Act as a virtual member of CoreWeave’s Compute Infrastructure, Fleet Operations, and Networking engineering teams, identifying opportunities for product enhancement and collaborating with engineers to implement your suggestions
  • Offer valuable insights on product features, functionality, and performance, contributing regularly to discussions about product strategy and architecture
  • Stay informed of the latest developments and trends in Kubernetes, cloud computing and infrastructure, sharing your thought leadership with customers and internal stakeholders
  • Lead the prototyping and initiation of research and development efforts for emerging products and solutions, delivering prototypes and key insights for internal consumption
  • Represent CoreWeave at conferences and industry events, with occasional travel as required

If you’re driven by innovation, thrilled by the possibilities of what specialized compute can enable, and eager to be part of a team that’s shaping the future, then CoreWeave is the place for you. Join us and let’s embark on this adventure together!

Qualifications

7+ years of proven experience as a Solutions Architect, Field Engineer, Infrastructure/Systems Engineer, or Technical Account Manager in Cloud Infrastructure, focusing on building or operating distributed systems or HPC/cloud services, with an expertise focused on bare‑metal compute infrastructure and large‑scale GPU cluster delivery.

  • Hands‑on experience bringing up, validating, and operating large GPU clusters—including bare‑metal node pxe boot, hardware health, fabric validation, and HPC acceptance/performance testing—and integrating bare‑metal with orchestration layers such as Kubernetes and Slurm
  • Deep expertise with modern rack‑scale GPU server hardware (e.g., NVIDIA HGX / GB200‑class systems), high‑speed interconnects (InfiniBand, NVLink), and the firmware/BMC/BIOS layer
  • Proven track record with building customer relationships, communicating clearly and the ability to break down complex technical concepts to both technical and non‑technical audiences
  • Expert‑level Linux system administration and command‑line troubleshooting, paired with strong networking fundamentals (routing, fabric topologies, TCP/IP)
  • B.S. in Computer Science or a related technical discipline, or equivalent experience

You’re curious about the latest and greatest technologies in the AI space.

You’re an expert in managing conflict and achieving mutually beneficial technical outcomes.

You love to help solve challenging technical problems.

Fluency in cloud computing concepts, architecture, and technologies with hands‑on experience in designing and implementing cloud solutions.

Preferred Qualifications
  • Experience operating security‑sensitive, air‑gapped, or otherwise locked‑down customer environments
  • Experience with scripting and automation related to bare‑metal provisioning, infrastructure validation, and lifecycle management (Python, Bash, Ansible, or similar)
  • Experience designing AI supercomputers from MEP designs
  • Experience delivering bare‑metal infrastructure at scale for large strategic customers or AI research labs
  • Experience with building solutions across multi‑cloud or hybrid environment
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Specialist Field Engineer - Compute Infrastructure
Senior Specialist Field Engineer - Compute Infrastructure

CoreWeave • Bellevue (WA)

On-site
USD 188,000 - 275,000
Medical, dental, and vision insurance
401(k) with employer match
Flexible PTO
Senior Specialist Field Engineer - Security
Senior Specialist Field Engineer - Security

CoreWeave • Livingston (NJ)

Hybrid
USD 165,000 - 220,000
Medical, dental, and vision insurance
401(k) with employer match
Flexible PTO
+2
Senior Specialist Field Engineer - Storage
Senior Specialist Field Engineer - Storage

Coreweave • San Francisco (CA)

On-site
USD 165,000 - 220,000
Medical, dental, and vision insurance – 100% paid
401(k) with employer match
Flexible PTO
+3
Senior Specialist Field Engineer - Security
Senior Specialist Field Engineer - Security

Coreweave • Livingston (NJ)

Hybrid
USD 165,000 - 220,000
Medical, dental, and vision insurance
401(k) with generous employer match
Flexible PTO
+1
Principal Solution Specialist, Core Services
Principal Solution Specialist, Core Services

CoreWeave • San Francisco (CA)

On-site
USD 198,000 - 264,000
Medical/Dental/Vision insurance
401(k) with company match
Tuition Reimbursement
+1
Principal Solution Specialist, Core Services
Principal Solution Specialist, Core Services

CoreWeave • Seattle (WA)

On-site
USD 198,000 - 264,000
Medical, dental, and vision insurance
401(k) with employer match
Paid Parental Leave
+1
Senior Specialist Field Engineer - Security
Senior Specialist Field Engineer - Security

CoreWeave • San Francisco (CA)

On-site
USD 165,000 - 220,000
Medical, dental, and vision insurance
Flexible PTO
Company-paid Life Insurance
+1
Principal Engineer, Cloud Infrastructure Services
Principal Engineer, Cloud Infrastructure Services

Weights & Biases • United States

Hybrid
USD 206,000 - 303,000
Medical, dental, and vision insurance
401(k) with employer match
Flexible PTO
+3
Senior Specialist Field Engineer - Security
Senior Specialist Field Engineer - Security

CoreWeave • Sunnyvale (CA)

Hybrid
USD 175,000 - 225,000
Health insurance
Life insurance
Disability coverage
+6
Principal Solution Specialist, Core Services
Principal Solution Specialist, Core Services

Coreweave • Washington

On-site
USD 198,000 - 264,000
Medical, dental, and vision insurance
Equity awards
401(k) with employer match
+4