Software Engineer - AI Infrastructure

Andromeda

San Francisco (CA)

Remote

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A cutting-edge AI infrastructure company is seeking an Infrastructure Product Engineer to design core platform components and build robust APIs. Candidates should have over 5 years of experience in infrastructure or backend engineering, with strong skills in Kubernetes and Python. This role offers autonomy and the opportunity to shape scalable AI infrastructure. Join us to help power the future of AI with the flexibility of a remote work environment.

Qualifications

  • 5+ years of experience in Infrastructure, Platform, or Backend Engineering roles.
  • Strong systems fundamentals including deep understanding of Linux, networking, storage, and distributed systems.
  • Proven expertise with Kubernetes, VMs, or bare-metal environments.

Responsibilities

  • Design and develop core platform components for infrastructure orchestration.
  • Build robust APIs and services that abstract diverse infrastructure types.
  • Collaborate with teams to define ownership boundaries between platform capabilities.

Skills

Infrastructure and Platform Engineering
Backend Engineering
Linux
Networking
Kubernetes
Python
Terraform
Ansible
Technical Communication

Tools

Kubernetes
Terraform
Ansible
Helm

Job description

Location: North America Remote / San Francisco

Employment Type: Full-Time

About Andromeda

Andromeda Cluster, founded by Nat Friedman and Daniel Gross, is on a mission to democratize access to cutting‑edge AI infrastructure previously reserved for hyperscalers. What began with a single managed cluster has quickly evolved into a global platform, connecting leading AI labs, data centers, and cloud providers. Our orchestration layer seamlessly routes training and inference jobs across the world, unlocking flexibility and efficiency in one of the fastest‑growing sectors on earth. Our long‑term vision is to establish a global marketplace for AI compute—powering AGI with the same fluidity as world financial markets.

The Role

As an Infrastructure Product Engineer, you will play a pivotal role in building the backbone of Andromeda’s platform. You'll transform complex, real‑world infrastructure challenges into scalable product capabilities that benefit our customers.

Positioned at the intersection of infrastructure and product engineering, this role is deeply technical and systems‑oriented, yet laser‑focused on building solutions with broad leverage.

What You’ll Do
  • Design and develop core platform components, including infrastructure orchestration, provisioning, and lifecycle management solutions.
  • Build robust APIs, services, and control planes that abstract over diverse infrastructure types (VMs, Kubernetes, bare metal, schedulers).
  • Translate customer usage patterns into product requirements, delivering impactful features and improvements.
  • Create automation and internal tooling to eliminate manual or ad‑hoc operational work.
  • Enhance reliability, performance, and observability at the platform level, emphasizing durable improvements over quick fixes.
  • Collaborate with peer teams to define clear ownership boundaries between platform capabilities and customer‑specific solutions.
  • Write clean, maintainable, and well‑documented code with a focus on long‑term sustainability.
  • Participate in technical design discussions and contribute to the architectural evolution of our platform.
What We’re Looking For
  • 5+ years of experience in Infrastructure, Platform, or Backend Engineering roles.
  • Strong systems fundamentals: deep understanding of Linux, networking, storage, and distributed systems.
  • Proven expertise with Kubernetes, VMs, or bare‑metal environments.
  • Advanced software engineering skills; capable of building production‑grade APIs and services (Python, Go, or similar).
  • Extensive experience with infrastructure as code and automation tools (Terraform, Ansible, Helm, etc.).
  • Demonstrated ability to navigate ambiguity and distill complex problems into clear, maintainable abstractions.
  • Product‑focused mindset: care about interfaces, defaults, reliability, and sustainable operations.
  • Excellent written and verbal communication skills; effective collaborator across engineering and product functions.
Nice to Have
  • Hands‑on experience with GPU or AI infrastructure.
  • Experience with control‑plane or orchestration systems.
  • Background spanning both infrastructure and application/backend engineering.
  • Experience architecting multi‑tenant systems.
  • Strong skills in technical writing and design documentation.
  • Early‑stage startup experience.
Why You’ll Love It Here

This is a true builder’s opportunity: you’ll have ownership and autonomy to shape our systems, engage directly with customers and providers, and lay the foundations for scalable, reliable AI infrastructure. Join us at Andromeda and help power the future of AI.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer - AI Infrastructure
Software Engineer - AI Infrastructure

Andromeda Cluster • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Ownership and autonomy in projects
Engagement with customers and providers
Opportunity to shape systems
Customer Reliability Engineer
Customer Reliability Engineer

Andromeda Cluster • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Andromeda • San Francisco (CA)

On-site
USD 150,000 - 200,000
Significant ownership and autonomy
Inclusive environment
Opportunity to shape AI infrastructure
Member of the Business Staff - Compute Markets
Member of the Business Staff - Compute Markets

Andromeda Cluster • San Francisco (CA)

Hybrid
USD 90,000 - 120,000
Competitive compensation
Meaningful equity
Comprehensive healthcare benefits
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Andromeda Cluster • San Francisco (CA)

Hybrid
USD 180,000 - 240,000
Infrastructure Manager
Infrastructure Manager

The Resume Database • San Francisco (CA)

On-site
USD 100,000 - 130,000
Competitive compensation
Meaningful equity
Comprehensive benefits including healthcare
+1
Compute Trader
Compute Trader

Andromeda • San Francisco (CA)

On-site
USD 100,000 - 140,000
Competitive compensation
Meaningful equity
Comprehensive healthcare benefits
+1
Member of the Technical Staff - Systems
Member of the Technical Staff - Systems

Andromeda • San Francisco (CA)

Hybrid
USD 190,000 - 260,000
Forward Deployed Engineer - SRE
Forward Deployed Engineer - SRE

Andromeda Cluster • San Francisco (CA)

Hybrid
USD 180,000 - 260,000
Health insurance
Equity
Unlimited PTO
+1
Technical Recruiter
Technical Recruiter

Andromeda • San Francisco (CA)

Hybrid
USD 120,000 - 180,000
Equity
Healthcare, dental, and vision
401(k)
+1