Get more replies from employers
Send a job-specific resume in minutes.
bespokelabs is looking for an Infrastructure Engineer in Mountain View, CA. This role involves owning the execution layer for RL environments and addressing complex systems challenges.
The ideal candidate should have a strong background in production systems, deep technical skills, and excellent communication abilities to collaborate with research teams.
The position offers competitive salary, equity, and health coverage along with the opportunity to work directly with leading AI research labs.
Bespoke Labs is an applied AI research lab pioneering data and RL environment curation for training and evaluating agents.
Recently, we curated Open Thoughts, one of the best open reasoning datasets used by multiple frontier labs, trained SOTA specialized models such as Bespoke-MiniChart-7B and Bespoke-MiniCheck, and built the environment infrastructure that frontier labs and enterprises use to make their agents reliable.
Bespoke is uniquely positioned to capture a large share of data and RL environment curation.
We're looking for an Infrastructure Engineer to own the execution layer beneath our RL environments: the systems that let an agent operate inside a realistic, multi-tool world coherently for hours or days.
This is a hard systems problem disguised as an AI job. As the tasks agents can complete keep lengthening, the environments that train them have to stay coherent across far longer horizons than anything that exists today. That means sandboxing and isolation you can trust, execution that's fast and cheap enough to run at training scale, and the ability to snapshot, restore, inspect, and branch a running environment instead of treating every rollout as one‑shot. You'll build the platform that makes all of this possible.
You will work closely with our research and data teams, and directly with frontier labs and enterprise customers, to turn environment designs into infrastructure that runs reliably in production.
Environment Execution & Sandboxing:
Performance & Scale
Environment Platform
Collaboration & Production Excellence
Systems & Infrastructure
Execution & Ownership
Collaboration & Communication
Experience with RL training or evaluation infrastructure, or the execution layer for agent rollouts.
Experience with checkpoint/snapshot‑restore systems, CRIU, or distributed state management.
Background in high‑throughput, low‑latency execution systems.
Contributions to widely‑used infrastructure, datasets, benchmarks, or open‑source systems.
Previous experience in a research engineering or infrastructure role at an AI or systems‑heavy company.
Location: Mountain View, CA
Compensation: Competitive salary and equity
Benefits: Health coverage, and the opportunity to work directly with the world's leading AI research labs