Senior Platform Engineer – AI/ML Infrastructure & Reliability

StratITech

San Francisco (CA)

On-site

USD 210,000 - 260,000

Full time

11 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Equity

Job summary

StratITech, a fast-growing San Francisco technology company, is seeking a senior platform engineer who loves building end-to-end infrastructure for AI workloads.

You will design Kubernetes-based platforms, automate deployments, and craft reusable tooling while collaborating closely with engineers, scientists, product teams, and leadership. Expect broad ownership in a small, autonomous team.

Qualifications

  • 8+ years in platform/infra engineering or related fields.
  • Evidence of owning end-to-end technical projects, not just operation.
  • Strong Linux fundamentals.
  • Terraform or IaC expertise.
  • Programming experience in Python or Go.
  • Experience building CI/CD and deployment automation.
  • Strong troubleshooting and distributed-systems understanding.
  • Ability to make architectural tradeoffs.

Responsibilities

  • Own platform and infrastructure projects from initial problem through production.
  • Design and build Kubernetes and container-based infrastructure.
  • Develop internal tooling, automation, and platform software.
  • Build reusable Infrastructure-as-Code and deployment systems.
  • Create CI/CD, GitOps, and developer self-service capabilities.
  • Solve technical problems across Linux, networking, storage, cloud, and distributed systems.
  • Build reliability and observability into systems from the beginning.
  • Partner with engineers, scientists, product teams, and leadership.
  • Own and evolve the systems you build after launch.

Skills

Kubernetes
Terraform
Go
Python
CI/CD
GitOps
Observability

Tools

ArgoCD
Helm
Docker

Job description

$210,000–$260,000 Base + Equity

About the Company

Our client is a fast-growing San Francisco technology company building an advanced AI platform on high-performance distributed infrastructure.

The engineering organization is small and highly technical. Engineers work closely with software teams, applied scientists, product, and leadership, with significant individual ownership and a short path from identifying a problem to deploying a production solution.

About the Role

We're looking for a senior platform engineer who is fundamentally a builder.

This is not a production operations role. We're looking for someone who has personally owned substantial technical projects end-to‑to‑end:

You’ll build foundational platform and infrastructure capabilities supporting advanced AI, data, and distributed workloads. Some projects will start with well-defined requirements; others will begin with an ambiguous technical problem that you will help define and solve.

The strongest candidates have experience in fast-growing startup environments, where engineers operate with significant autonomy, responsibilities are broad, and building from scratch is part of the job.

What You'll Do
  • Own platform and infrastructure projects from initial problem through production
  • Design and build Kubernetes and container-based infrastructure
  • Develop internal tooling, automation, and platform software
  • Build reusable Infrastructure-as-Code and deployment systems
  • Create CI/CD, GitOps, and developer self-service capabilities
  • Solve technical problems across Linux, networking, storage, cloud, and distributed systems
  • Build reliability and observability into systems from the beginning
  • Partner directly with engineers, scientists, product teams, and technical leadership
  • Own and evolve the systems you build after launch
What We're Looking For
  • 8+ years in platform engineering, infrastructure engineering, production engineering, SRE, DevOps, systems engineering, or related software engineering
  • Evidence of building or substantially redesigning systems—not simply operating them
  • Strong Linux and systems fundamentals
  • Terraform or comparable Infrastructure-as-Code expertise
  • Python, Go, or comparable programming/automation experience
  • Experience building CI/CD, deployment automation, or developer-platform capabilities
  • Strong troubleshooting and distributed-systems fundamentals
  • Ability to make architectural tradeoffs across performance, reliability, security, cost, and complexity
  • Comfort operating independently when requirements are incomplete or changing
These Skills Are a Plus
  • Fast-growing startup experience
  • Founding or early infrastructure/platform engineering experience
  • GPU, NVIDIA, or HPC infrastructure
  • AI/ML training or inference infrastructure
  • Distributed compute or high-throughput systems
  • Kafka, Spark, Airflow, Ray, or Databricks
  • Advanced networking or distributed storage
  • GitOps, ArgoCD, Helm, or service mesh technologies
You’ll Thrive Here If

You enjoy building more than maintaining. You want meaningful technical ownership, are comfortable moving outside a narrow specialty, and don't need someone assigning your next ticket.

You can take an ambiguous problem, determine what needs to be built, evaluate the tradeoffs, design the solution, implement it, and take responsibility for the result.

Why Join
  • Build foundational systems rather than simply maintain existing infrastructure
  • Own technically meaningful projects from concept through production
  • Work on challenging AI, distributed-systems, data, and compute problems
  • Join a small engineering organization where individual contributions matter
  • Work directly with highly technical engineers, scientists, and leadership
  • $210,000–$260,000 base salary plus equity

Work Authorization: Candidates must be currently authorized to work in the United States. This position is not eligible for new or future employer-sponsored work authorization.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Technical Lead – Data Platform
Technical Lead – Data Platform

StratITech • San Francisco (CA)

On-site
USD 240,000 - 270,000
Equity
Benefits
Senior Software Engineer - Infrastructure/Platform
Senior Software Engineer - Infrastructure/Platform

Thomas Talent Network • San Francisco (CA)

On-site
USD 250,000 - 350,000
Platform Engineer
Platform Engineer

Goliath Partners Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 230,000 - 285,000
Medical, dental & vision
401(k)
Visa sponsorship
+1
Platform Engineer
Platform Engineer

Goliath Partners • San Francisco (CA)

Hybrid
USD 230,000 - 285,000
Med, dental, vision coverage
401(k)
Visa sponsorship
+2
Founding Platform Engineer - FinTech
Founding Platform Engineer - FinTech

Leadervest • New York (NY)

On-site
USD 180,000 - 250,000
Founding equity
Senior Site Reliability Engineer
Senior Site Reliability Engineer

The Recruiting Guy • Arlington (VA)

On-site
USD 175,000 - 250,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

The Recruiting Guy • Washington

On-site
USD 175,000 - 250,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

The Recruiting Guy • San Francisco (CA)

On-site
USD 175,000 - 250,000
Senior Software Engineer, Backend
Senior Software Engineer, Backend

Recruiting from Scratch • San Francisco (CA)

On-site
USD 250,000 - 300,000
Competitive equity package
Profit share bonus
Director of Infrastructure Engineering
Director of Infrastructure Engineering

Appsierra Group • United States

On-site
USD 350,000 - 500,000
Equity compensation eligibility
Performance-based bonuses
Health insurance reimbursement up to 1
+3