Infrastructure Engineer

Nebula

El Segundo (CA)

Hybrid

USD 100,000 - 250,000

Full time

17 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Equity participation
Health, dental, vision
401(k) with matching
Nearby housing bonus
On-site work

Job summary

Nebula is seeking an Infrastructure Engineer to run and harden the company’s self-hosted Linux, Kubernetes, networking, storage, and security stack on hardware we own. You will help make the estate reproducible via version control, plan capacity, and manage GPU AI compute workloads on-site in El Segundo.

Experience with Linux, Kubernetes in production, backups, and security discipline is required. GPU infra, Ansible or Terraform, and self-hosted CI tooling are a plus.

Qualifications

  • Degree in computer engineering or computer science, or equivalent experience.
  • Hands-on Linux administration on owned machines: systemd, networking, storage and filesystems.
  • Kubernetes in production: workloads, storage, ingress, upgrades.
  • Fluent with routing, firewalls, DNS, VPNs, TLS and certificates.
  • Backups proven by restores; security discipline expected.
  • Ready to put a hand-built estate into code and diagnose outages.

Responsibilities

  • Bare-metal Linux hosts provisioning, systemd, disks/RAID, encryption at rest, out-of-band management.
  • Kubernetes cluster: scheduling, ingress and TLS, PVs, network policies, upgrades/rollbacks.
  • Network, access and identity: routing, firewalling, VPN, DNS, certs, mTLS, SSO.
  • Storage and backups: capacity planning, retention, offsite copies, restore drills.
  • Monitoring and incident response: metrics, logs, alerts, on-call, postmortems.
  • Security: hardening, patching, secrets, key rotation, least privilege.
  • Infrastructure as code: version control, self-hosted SCM, registry and CI runners.
  • Capacity/hardware planning: hosts, storage, network, GPU AI workloads.

Skills

Linux admin
Kubernetes prod
Networking
Storage mgmt
Security discipline
IaC tooling
GPU infra planning

Education

Bachelors in CS/EE or equivalent

Tools

Ansible
Terraform
Git
CI runners

Job description

What to Expect

Your role is to run and harden the infrastructure our own software and our AI systems run on: the servers, the network, the storage and the Kubernetes cluster behind them, all self-hosted on hardware we own.

That is bare-metal Linux and a cluster on top of it, the network and identity by which people and machines reach the estate, backups proved by restores, monitoring and incident response, and security as a daily habit.

You will take an estate built by hand and make it reproducible from version control, plan capacity and hardware ahead of demand, including the GPU compute our AI workloads run on, and extend the agents that carry routine operational work.

We are looking for infrastructure engineers with experience and interest in running Linux and Kubernetes on hardware they own, networking and storage, backups and the restores that prove them, and security and automation.

Every role works with our AI systems daily. What can be deterministic, must be. You need no AI background; we prefer people without one. We hire for your knowledge and experience in the field first, so you can steer the ship; the tooling is a learning curve we expect you to take on.

We work on-site in El Segundo. This is hard work, but you will be rewarded with equity in a company we believe will become one of the world's most valuable.

What You'll Do
  • Bare-metal Linux hosts: provisioning, systemd, disks and RAID, encryption at rest, out-of-band management, racking
  • The Kubernetes cluster: scheduling, ingress and TLS, persistent volumes, network policies, upgrades and rollbacks
  • Network, access and identity: routing and firewalling, VPN, DNS, certificates, mutual TLS, single sign-on
  • Storage and backups: capacity layout, retention design, offsite copies, and restore drills actually run
  • Monitoring and incident response: metrics, structured logs, alert routing, on-call, postmortem follow-through
  • Security: hardening and patching, secrets management, key rotation, least privilege, external exposure review
  • Infrastructure as code: configuration in version control, self-hosted source control, registry and CI runners
  • Capacity and hardware planning: hosts, storage, network, and the GPU compute AI workloads need
What You'll Bring
  • A degree in computer engineering or computer science, a systems administration background, or equivalent experience
  • Hands-on Linux administration on machines you were responsible for: systemd, networking, storage and filesystems
  • Kubernetes in production: workloads and storage you sized, ingress you configured, upgrades you have run
  • Fluent with routing, firewalls, DNS, VPNs, TLS and certificates when something is unreachable
  • Backups proved by restores you have performed, and security discipline held to without being asked
  • Ready to put a hand-built estate into code, and to take an outage to its cause
  • Bonus: GPU and machine learning infrastructure, Ansible or Terraform on a real estate, or serious self-hosting
Expected Compensation

$100k to $250k base salary, plus equity

  • Equity participation in a high-growth startup
  • Comprehensive health, dental, and vision insurance
  • 401(k) with company matching
  • Bonus for living within 5 miles of our El Segundo facility
  • On-site work with rare work-from-home exceptions
  • Merit-based organization where contribution drives reward
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Infrastructure
Software Engineer, Infrastructure

descript • United States

On-site
USD 220,000 - 292,000
Equity
Competitive benefits
Software Engineer
Software Engineer

Nebula • El Segundo (CA)

Hybrid
USD 100,000 - 250,000
Equity
Health insurance
Dental insurance
+3
Founding Engineer - ML Infrastructure
Founding Engineer - ML Infrastructure

uRun • San Francisco (CA)

On-site
USD 120,000 - 160,000
Health, dental, and vision
401(k)
Paid time off
+2
Founding Infrastructure Engineer
Founding Infrastructure Engineer

Matterhaul Inc. • San Francisco (CA)

On-site
USD 200,000 - 260,000
Equity options
Hardware + AI coding budget
Real office in SF
Data Platform Engineer
Data Platform Engineer

Nebula • El Segundo (CA)

Hybrid
USD 100,000 - 250,000
Equity participation
Health, dental and vision insurance
401(k) with company matching
+1
Member of Technical Staff, Infrastructure Engineer
Member of Technical Staff, Infrastructure Engineer

Edison Scientific • San Francisco (CA)

On-site
USD 175,000 - 240,000
Health care coverage
Equity offered
Parental leave
Staff Platform Engineer – Infrastructure
Staff Platform Engineer – Infrastructure

Recruiting from Scratch • United States

On-site
USD 200,000 - 300,000
20 days PTO annually
100% covered health insurance
401(k) with employer contribution
+1
Agent Infrastructure Engineer
Agent Infrastructure Engineer

Nebula • El Segundo (CA)

Hybrid
USD 100,000 - 250,000
Equity participation in a high-growth 
Comprehensive health, dental, and VIs
401(k) with company matching
+3
Member of Technical Staff, Platform - AI Infrastructure
Member of Technical Staff, Platform - AI Infrastructure

Hamilton Barnes Associates Limited • United States

On-site
USD 213,000 - 288,000
Equity
Health care
Platform Engineer
Platform Engineer

Harper • San Francisco (CA)

On-site
USD 140,000 - 280,000
Uber commuter benefits
Meals provided (breakfast, lunch, and/
Snacks, drinks and coffee daily
+2