DevOps Engineer, Infrastructure & Platforms

Ricursive Intelligence

Palo Alto (CA)

On-site

USD 150,000 - 210,000

Full time

7 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Ricursive Intelligence in Palo Alto is seeking an experienced Infrastructure/Platform Engineer to own IaC, CI/CD, and cloud platforms for a fast-growing AI research lab. You will design and operate self-managed cloud infrastructure, implement observability, and collaborate with PD engineers to keep research pipelines reliable, secure, and scalable in a mission-driven startup environment.

You will work with Terraform, GitHub Actions, Kubernetes, and GCP, ensuring reproducible deployments and

Qualifications

  • BS in CS, CE, EE or equivalent practical experience.
  • 4+ years of hands-on infrastructure/platform engineering, with ownership of systems others depend on.
  • Owned production IaC architecture including maintenance and new features, not just consumed modules.
  • Designed and implemented production CI/CD pipelines that build and ship reproducible artifacts, with attention to performance, scalability, and security.
  • Prior experience with observability tooling for monitoring pipeline health as well as application-level metrics.

Responsibilities

  • Design, build out and extend our self-managed cloud platform with Terraform, setting the IaC patterns the rest of the team builds on.
  • Own our platform deployments spanning various environments day to day, including performance, security, reliability, scalability, and adapting to evolving business requirements.
  • Design, implement, and operate CI/CD pipelines in GitHub Actions for mission-critical repositories, working within strict security and deployment restrictions.
  • Architect scalable AI tooling and developer experience workflows across multiple distinct user environments, working closely with our physical design (PD) engineers.
  • Build out the observability stack with the team, covering pipeline health, application-level metrics, and ML workloads.

Skills

IaC
CI/CD
Cloud platforms
Observability
Kubernetes
GCP
Terraform
GitHub Actions
Security awareness
Automation

Education

BS in CS/CE/EE or equivalent

Tools

Terraform
GitHub Actions
Kubernetes
GCP
CI/CD tooling

Job description

Ricursive Intelligence is a frontier AI lab building self-improving systems, starting with chip design. We are reinventing chip development and closing the loop between AI and the hardware that fuels it, recursively accelerating the path to artificial superintelligence. Backed by $335M from Sequoia, Lightspeed, DST, and NVIDIA Ventures, we are a growing, fast-paced team where every hire shapes the work.

The company has unmatched talent density, including IMO, IPHO, and IOAA gold medalists, pioneers who made prior breakthroughs in chip design: AlphaChip (Nature 2021), ePlace (DAC Best Paper Nominee 2014), RL-CCD (DAC Best Paper 2023), INSTA (DAC Best Paper 2025), and C3PO (ASP-DAC Best Paper 2026), chip leads for Apple Silicon (multiple generations of iPhone and iPad) and Google (TPU, OpenTitan), and top researchers and engineers from Anthropic, Google DeepMind, Stanford, and MIT.

ABOUT THE ROLE

Ricursive's research runs on infrastructure that has to keep pace with the research itself: ML training and evaluation workloads, EDA tool flows, and a fast-growing team that needs everything from cloud environments to developer workflows to just work. This role owns that foundation — the pipelines, platforms, and systems that let a small team move like a much larger one.

You will own our infrastructure-as-code (IaC), continuous integration & deployment (CI/CD), and cloud platform end-to-end that keeps the lab running day to day. Additionally, you will be collaborating closely with the team to build out the right observability stack for their needs while working around environment security limitations.

WHAT YOU WILL DO
  • Design, build out and extend our self-managed cloud platform with Terraform, setting the IaC patterns the rest of the team builds on.
  • Own our platform deployments spanning various environments day to day, including performance, security, reliability, scalability, and adapting to evolving business requirements.
  • Design, implement, and operate CI/CD pipelines in GitHub Actions for mission-critical repositories, working within strict security and deployment restrictions.
  • Architect scalable AI tooling and developer experience workflows across multiple distinct user environments, working closely with our physical design (PD) engineers.
  • Build out the observability stack with the team, covering pipeline health, application-level metrics, and ML workloads.
MINIMUM QUALIFICATIONS
  • BS in CS, CE, EE, or a closely related technical field, or equivalent practical experience.
  • 4+ years of hands‑on infrastructure/platform engineering, with ownership of systems others depend on.
  • Owned production IaC architecture including maintenance and new features, not just consumed modules.
  • Designed and implemented production CI/CD pipelines that build and ship reproducible artifacts, with attention to performance, scalability, and security.
  • Prior experience with observability tooling for monitoring pipeline health as well as application-level metrics.
PREFERRED QUALIFICATIONS
  • Hands‑on experience with GCP, Kubernetes, and GitHub Actions, including custom runner setups.
  • Experience running ML infrastructure for training and evaluation workloads, including GPU/TPU compute.
  • Familiarity with LLM observability tooling: tracing, evaluations, and cost & latency monitoring.
  • Security depth: dependency supply‑chain hardening, OIDC-based auth, least‑privilege secrets, and compliance work such as SOC 2 or penetration testing.
  • Early‑stage startup experience: built infrastructure from zero or near‑zero.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Platform & IaC Engineer for AI Hardware Infra
Platform & IaC Engineer for AI Hardware Infra

Ricursive Intelligence • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Platform Engineer - AI/ML Infrastructure (Kubernetes & Terraform)
Platform Engineer - AI/ML Infrastructure (Kubernetes & Terraform)

Madrona Venture Labs • United States

On-site
USD 180,000 - 260,000
Member of Technical Staff - SWE Infrastructure
Member of Technical Staff - SWE Infrastructure

Ricursive Intelligence • Palo Alto (CA)

On-site
USD 100,000 - 140,000
Member of Technical Staff - Compute Platform
Member of Technical Staff - Compute Platform

Prime Intellect AI • San Francisco (CA)

On-site
USD 150,000 - 230,000
Member of Technical Staff - Compute Platform
Member of Technical Staff - Compute Platform

Prime Intellect • San Francisco (CA)

On-site
USD 140,000 - 190,000
Hybrid work model
Open-source contributions encouraged
Member of Technical Staff [Platform]
Member of Technical Staff [Platform]

NeoCognition Inc. • Palo Alto (CA)

On-site
USD 120,000 - 160,000
DevOps Engineer (Founding Team)
DevOps Engineer (Founding Team)

Fabrion • San Francisco (CA)

On-site
USD 120,000 - 160,000
Competitive salary
Meaningful equity
Member of Technical Staff, Infrastructure
Member of Technical Staff, Infrastructure

Psi • Boston (MA), Northern (KY)

On-site
USD 180,000 - 260,000
Meaningful equity
Competitive compensation
Benefits
Member of Technical Staff - Sandbox Platform
Member of Technical Staff - Sandbox Platform

Prime Intellect AI • San Francisco (CA)

Hybrid
USD 150,000 - 300,000
Cash compensation range: $150-300k
Flexible work arrangement (SF office +
Full visa sponsorship and relocation
+1
AI Engineer
AI Engineer

Pinpoint Global Communications • United States

On-site
USD 120,000 - 180,000