Senior AI Platform SWE - PlayerZero

HireOTS

Atlanta (GA)

On-site

USD 120,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A stealth-stage AI infrastructure startup is seeking a Platform Engineer to build a self-healing system. In this role, you'll design scalable cloud architectures and manage complex data pipelines to automate software defect resolution. Ideal candidates have extensive experience in distributed systems and a proactive approach to infrastructure ownership.

Qualifications

  • 5-10+ years of experience building and maintaining distributed systems in production.
  • Deep understanding of consensus protocols, sharding, replication, and backpressure.
  • Experience managing high-volume, multi-tenant data pipelines.

Responsibilities

  • Design and build cloud-native architectures that support thousands of microservices.
  • Implement low-latency indexes for semantic queries across massive codebases.
  • Automate LLM-agent orchestration for concurrent autonomous fix bots.

Skills

Distributed Systems
Cloud-native Architectures
Data Pipelines
Infrastructure as Code

Tools

Terraform
Pulumi
Argo
GitHub Actions
Buildkite

Job description

A stealth-stage AI infrastructure startup is building a self-healing system for software that automates defect resolution and development. Our platform is used by engineering and support teams to:

  • Autonomously debug problems in software (technical support)

  • Fix issues directly in code

  • Prevent these problems from recurring

The company is backed by leading investors including Foundation Capital, WndrCo, and Green Bay Ventures — along with prominent operators such as Matei Zaharia, Drew Houston, Dylan Field, Guillermo Rauch, and others.

We believe that as software development accelerates, engineering and support teams face mounting challenges maintaining quality and reliability. We see this as a rare opportunity to reinvent how modern software is supported — with intelligent, automated infrastructure.

About the Role

We’re searching for a Platform Engineer who lives and breathes distributed systems. You’ll help build the foundation that synchronizes petabytes of data, indexes billions of lines of code, and coordinates fleets of AI agents working in parallel. If terms like “five-nines,” “exact-once,” or “sub-second latency” light you up — you’ll thrive here.

In this role, you will:
  • Design cloud-native architectures that elastically scale to support thousands of microservices and GPU workers.

  • Build high-throughput data planes for log ingestion, change-data-capture (CDC), and real-time feature stores powering our self-healing system.

  • Implement ultra-low-latency indexes (vector, inverted, graph) to support semantic queries across massive codebases.

  • Synchronize state across clusters and regions for global enterprise users — ensuring consistent, fast results regardless of data scale.

  • Automate LLM-agent orchestration so hundreds of autonomous fix bots can work concurrently without collisions.

  • Harden reliability and security with chaos engineering, live migrations, and deep protections for customer data.

  • Partner closely with ML researchers and product leads to translate novel ideas into hardened infrastructure.

You might thrive in this role if you:
  • Have 5–10+ years of experience building and maintaining distributed systems in production.

  • Understand consensus protocols, sharding, replication, and backpressure inside and out.

  • Have successfully managed high-volume, multi-tenant data pipelines in real-time environments.

  • Are comfortable navigating and debugging massive codebases and infra at scale.

  • Write infra-as-code (Terraform / Pulumi) and automate with tools like Argo, GitHub Actions, or Buildkite.

  • Take full ownership: you architect it, you build it, you run it.

Let me know if you'd like this version made more company-facing (for job boards) or stealth investor-facing (for fundraising decks).

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Software Engineer - Infrastructure / Platform (AI Startup)
Senior Software Engineer - Infrastructure / Platform (AI Startup)

Mintstage • California (MO)

On-site
USD 140,000 - 200,000
Senior Software Engineer
Senior Software Engineer

Xcede • San Francisco (CA)

On-site
USD 180,000 - 260,000
Principal Software Engineer - Platform
Principal Software Engineer - Platform

Atlan • United States

Remote
USD 180,000 - 250,000
Software Engineer, Platform
Software Engineer, Platform

Recruiting from Scratch • San Francisco (CA)

On-site
USD 140,000 - 210,000
Bi-annual bonuses
Relocation assistance
Housing stipend
+5
Software Engineer (AI Platform)
Software Engineer (AI Platform)

Superblocks • New York (NY)

On-site
USD 175,000 - 225,000
Generous equity package
Comprehensive benefits
Member of Technical Staff - Backend & Infrastructure
Member of Technical Staff - Backend & Infrastructure

janitorAI • San Francisco (CA)

On-site
USD 120,000 - 160,000
Platform Engineer
Platform Engineer

Harper • San Francisco (CA)

On-site
USD 140,000 - 280,000
Uber commuter benefits
Meals provided (breakfast, lunch, and/
Snacks, drinks and coffee daily
+2
Founding Software Engineer
Founding Software Engineer

Aurrum Services • New York (NY)

Hybrid
USD 150,000 - 210,000
AI Platform Engineer
AI Platform Engineer

Glint Tech Solutions • San Francisco (CA)

On-site
USD 180,000 - 240,000
Software Engineer (Agent Infra)
Software Engineer (Agent Infra)

Hinoki • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000