Une candidature sur mesure pour ce poste — un CV personnalisé et une lettre de motivation qui correspondent directement à l’offre.
White Circle, an AI safety company, is hiring a ML Infrastructure Engineer to design and operate scalable RL and post‑training pipelines. You will build data control systems for rollouts, replay, evaluation, and policy updates, and tune training and inference end‑to‑end for throughput across GPUs, networking, and storage.
Expect work on experiment tooling, dashboards, and cost visibility. You will relocate to Paris on a hybrid basis and contribute to agent infrastructure, runtime sandboxes, and
White Circle, an AI Safety company building the policy enforcement and optimization layer for AI systems. Backed by $11M from senior leaders at OpenAI, Anthropic, HuggingFace, Mistral, and DeepMind, White Circle processes 100M+ API calls monthly and runs its own LLMs in production.
You will Build scalable RL and post-training pipelines, including smoke tuning runs for quality testing and ablations. Design data control systems for rollouts, replay, filtering, evaluation, and policy updates. Tune training and inference end-to-end for throughput: networking, memory, scheduling, data loading, storage, checkpointing, I/O. Build infrastructure for model iteration (experiment runs, artifacts, evals, dashboards, reproducibility, cost visibility) and inference infrastructure for post-training and eval loops. Build agentic development environments: coding-agent harnesses, tool integrations, runtime sandboxes, multi-agent orchestration.