AI Infrastructure Engineer: High-Scale Sandboxing & RL

Socket.dev

Palo Alto (CA)

On-site

USD 210,000 - 260,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Healthcare coverage
Relocation support
Wellness programs

Job summary

Mistral is seeking a seasoned Systems/Distributed Infrastructure Engineer to lead end-to-end execution and data infrastructure powering its agentic models and coding assistants. You will help design scalable synthetic data generation, ultra-fast training environments, and high-throughput execution engines across hybrid and multi-cloud clusters.

You will tackle challenges from 1 million concurrent sandboxes to optimize training codebases, pipelines, and orchestration, working with researchers to

Qualifications

  • 4+ years in Systems Engineering/Distributed Systems/Cloud Infra or MLOps.
  • Experience building high-throughput data processing pipelines and generation pipelines for large-scale datasets.
  • Strong proficiency in Python, Go, C++ or Rust; experience profiling and optimizing high-performance ML/backend code.

Responsibilities

  • Design, deploy and operate high-throughput sandboxing for untrusted code across 1M+ environments.
  • Architect data generation pipelines for synthetic code, agent trajectories, rollouts and self-play data.
  • Optimize agent training codebases and distributed runtimes (Python, Go, C++, Rust) to improve GPU utilization and reduce rollout overhead.
  • Reduce sandbox cold-start times with container warm pools and efficient image delivery across hybrid clusters.
  • Develop Kubernetes-native controllers and queuing systems for diverse hardware fleets.
  • Ensure multi-tenant isolation and secure boundary enforcement across runtimes and networks.

Skills

Distributed Systems
Cloud Infra
MLOps
Kubernetes
Container Tech
Python
Go/C++/Rust
4+ years experience

Tools

Kubernetes
Docker
CRIU
Firecracker
gVisor

Job description

Mistral is seeking a seasoned Systems/Distributed Infrastructure Engineer to lead end-to-end execution and data infrastructure powering its agentic models and coding assistants. You will help design scalable synthetic data generation, ultra-fast training environments, and high-throughput execution engines across hybrid and multi-cloud clusters.

You will tackle challenges from 1 million concurrent sandboxes to optimize training codebases, pipelines, and orchestration, working with researchers to

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RL Environment Infrastructure Engineer — Sandbox & Scale
RL Environment Infrastructure Engineer — Sandbox & Scale

Commergence • Colorado

Hybrid
USD 150,000 - 210,000
RL Sandbox & Infra Engineer - Scalable Run-Time Platform
RL Sandbox & Infra Engineer - Scalable Run-Time Platform

Commergence • Fremont (CA)

Hybrid
USD 120,000 - 180,000
AI Infrastructure Engineer — Kubernetes, Linux & Cloud
AI Infrastructure Engineer — Kubernetes, Linux & Cloud

Lindus Health • Palo Alto (CA)

On-site
USD 140,000 - 200,000
Research Engineer, Code Agents Infra
Research Engineer, Code Agents Infra

Socket.dev • Palo Alto (CA)

On-site
USD 210,000 - 260,000
Healthcare coverage
Relocation support
Wellness programs
ML Platform Engineer — Scalable GPU & Kubernetes
ML Platform Engineer — Scalable GPU & Kubernetes

Socket.dev • Palo Alto (CA)

On-site
USD 180,000 - 280,000
Healthcare coverage
Relocation support
Retirement plans
+2
AI Lab Tech Engineer
AI Lab Tech Engineer

Commergence • Fremont (CA)

Hybrid
USD 120,000 - 180,000
Senior SRE: Scalable AI Platform & HPC
Senior SRE: Scalable AI Platform & HPC

Mistral Ai • New York (NY)

On-site
USD 150,000 - 180,000
AI Lab Tech Engineer
AI Lab Tech Engineer

Commergence • Colorado

Hybrid
USD 150,000 - 210,000
Senior Cloud Platform SRE for Scalable AI Systems
Senior Cloud Platform SRE for Scalable AI Systems

Mistral AI • Germany (OH)

On-site
USD 130,000 - 210,000
Healthcare coverage
Relocation support
Retirement plans
+2
ML Platform Engineer: GPU Orchestration & Scale
ML Platform Engineer: GPU Orchestration & Scale

Mistral • Palo Alto (CA)

On-site
USD 180,000 - 280,000
Healthcare coverage
Parental leave
Retirement plans
+3