Principal Data Center Infrastructure Software Engineer

Designworks Talent

Bellevue (WA)

Hybrid

USD 150,000 - 210,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Designworks Talent is seeking a Data Center Infrastructure Software Engineer to join our AI infrastructure group in Bellevue, WA. This hybrid role focuses on designing, deploying, and automating data center software that powers large-scale AI workloads, from server provisioning to GPU clusters and production inference.

Collaborate on infrastructure-as-code, Kubernetes deployments, and performance optimization in a lean, AI-native environment with cross-functional teams.

Qualifications

  • 5+ years designing, building, or operating large-scale Linux-based infrastructure.
  • Hands-on experience with Kubernetes, containerization, and distributed systems in production environments.
  • Experience with infrastructure-as-code and automation tools such as Terraform, Ansible.

Responsibilities

  • Develop infrastructure-as-code, automation, and provisioning systems for compute, networking, and storage.
  • Deploy and optimize Kubernetes, container, and distributed computing platforms.
  • Optimize GPU, networking, storage, and system performance for large-scale AI workloads.
  • Troubleshoot complex issues across hardware, operating systems, networking, storage, and software stacks.
  • Build reliability, observability, and operational excellence practices for mission-critical infrastructure.

Skills

Kubernetes
Linux
Distributed systems
Automation
Containerization
Terraform
Ansible

Tools

Terraform
Ansible

Job description

Data Center Infrastructure Software Engineer

Location: Hybrid | Bellevue, WA (downtown)
Titles:
Senior | Staff | Principal (multiple roles available)

Build the Data Center Software Infrastructure
About the Opportunity

A well-funded, rapidly growing AI infrastructure company is building a next-generation cloud platform designed to power the full lifecycle of artificial intelligence. The organization is developing a comprehensive AI infrastructure, platform, and services portfolio that supports the full spectrum of AI workloads including large-scale compute, model training, fine-tuning, inference, and emerging agentic AI applications.

Backed by significant long-term investment, the company combines the speed, ownership, and innovation of a startup with the stability and resources of an established parent organization. Engineering teams are intentionally lean, highly collaborative, and AI-native, leveraging modern tooling and automation to build infrastructure capable of supporting the industry's most demanding AI workloads.

We're seeking Data Center Software Engineers to lead the design, development, configuration, and automation of AI infrastructure clusters.

The Opportunity

Your responsibility begins once servers and racks are installed in the data center and extends through software deployment, networking, configuration, cluster bring-up, and automation, ensuring the platform is fully operational and ready for customer workloads.

What You'll Do
  • Develop infrastructure-as-code, automation, and provisioning systems for compute, networking, and storage.

  • Deploy and optimize Kubernetes, container, and distributed computing platforms.

  • Optimize GPU, networking, storage, and system performance for large-scale AI workloads.

  • Troubleshoot complex issues across hardware, operating systems, networking, storage, and software stacks.

  • Build reliability, observability, and operational excellence practices for mission-critical infrastructure.

What We're Looking For
  • 5+ years of experience designing, building, or operating large-scale Linux-based infrastructure.

  • Hands-on experience with Kubernetes, containerization, and distributed systems in production environments.

  • Experience with infrastructure-as-code and automation tools such as Terraform, Ansible, or similar frameworks.

  • Strong experience operating cloud or datacenter-scale infrastructure.

Preferred Qualifications
  • Experience with bare-metal provisioning and hardware lifecycle management platforms (e.g., MAAS, Ironic, xCAT, xCAT, or similar).

  • Experience with IPMI, Redfish, PXE boot, and automated operating system deployment at scale.

  • Experience managing GPU clusters in datacenter or cloud environments.

Compensation
  • Competitive base pay for Bellevue market

  • Certain roles are eligible for additional rewards, including merit increases, annual bonus, and long term incentives. These awards are allocated based on individual performance

  • U.S. based employees have access to medical, dental, and vision insurance, a 401(k) plan and company match, employees also receive per calendar year, paid holidays.

Location
  • Hybrid role based in the Bellevue, WA area.

  • Approximately three days per week in the office.

  • Candidates elsewhere in the U.S. who are open to relocation are encouraged to apply.

  • U.S. work authorization is required. Visa sponsorship is not currently available.

Why Join?
  • Ground-floor opportunity: you'll be among the earliest engineers on the team, directly shaping architecture, tooling, and culture.

  • Work directly on cutting-edge AI infrastructure at real scale — from data center design through GPU clusters to production inference.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

VP - AI Infrastructure Engineering
VP - AI Infrastructure Engineering

Designworks Talent • Bellevue (WA)

Hybrid
USD 300,000 - 520,000
Data Center Operations and Maintenance Engineer
Data Center Operations and Maintenance Engineer

Designworks Talent • Bellevue (WA)

Hybrid
USD 130,000 - 190,000
Medical insurance
Dental and vision insurance
401(k) with company match
+1
Data Center Operations and Maintenance Engineering Leader
Data Center Operations and Maintenance Engineering Leader

Designworks Talent • Bellevue (WA)

Hybrid
USD 180,000 - 240,000
Health insurance
Vision coverage
401(k) plan
Hardware Design Engineer
Hardware Design Engineer

Designworks Talent • Bellevue (WA)

Hybrid
USD 180,000 - 240,000
Medical, dental, vision insurance
401(k) with company match
Paid holidays
Principal Data Center Infra & AI Platform Engineer
Principal Data Center Infra & AI Platform Engineer

Designworks Talent • Bellevue (WA)

Hybrid
USD 150,000 - 210,000
Senior Software Engineer, Backend
Senior Software Engineer, Backend

Recruiting from Scratch • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive equity package
Profit share bonus
Data Center Infrastructure Architect
Data Center Infrastructure Architect

OpenAI • San Francisco (CA)

On-site
USD 200,000 - 280,000
Data Center Infrastructure Architect
Data Center Infrastructure Architect

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
VP – AI Infrastructure Engineering
VP – AI Infrastructure Engineering

Jobtailor • Bellevue (WA)

On-site
USD 200,000 - 350,000
Infrastructure Design Engineer
Infrastructure Design Engineer

Togetherai • San Francisco (CA)

On-site
USD 210,000 - 250,000
Startup equity
Health insurance
Flexible remote work