Member of Technical Staff — Infrastructure

Remanence

Paris (TX)

Hybrid

USD 124,000 - 186,000

Full time

11 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Visa sponsorship
Relocation support
Hybrid work setup
Equity

Job summary

Remanence in Europe seeks a Member of Technical Staff - Infrastructure to own the compute and execution platform for training, inference and evaluation. You will ensure GPU resources are productive and research workflows are easy to operate and debug.

You will build and run GPU clusters, scheduling, networking and storage for distributed ML workloads; develop the execution platform for many concurrent environments; create data and artifact pipelines and improve observability, latency and

Qualifications

  • Strong systems fundamentals and careful reasoning about concurrency, resources and failure recovery.
  • Ability to make complex infrastructure understandable and dependable for users.
  • Prior ML infrastructure experience is valuable.

Responsibilities

  • Build and operate GPU clusters, job scheduling, networking and storage for distributed ML workloads.
  • Develop the execution platform for large numbers of concurrent environments, including sandbox isolation, resource limits, retries, and state recovery.
  • Build reliable data and artifact pipelines for datasets, trajectories, model weights, and checkpoints.
  • Own platform observability and recovery; improve capacity allocation, startup latency, and reliability through measurable changes.
  • Provide researchers with reproducible environments and simple tools to launch, inspect, and debug experiments.

Skills

Systems fundamentals
Concurrency
Distributed ML
Observability

Tools

Linux
Docker
Kubernetes
Slurm
Terraform
Python
Go
Ray
PostgreSQL
Prometheus
Grafana
OpenTelemetry
NVIDIA DCGM

Job description

Member of Technical Staff - Infrastructure

Remanence is pioneering the next era of enterprise AI by building intelligent systems that learn continuously from real-world execution. We transform complex enterprise workflows and business context into dynamic, interactive environments where AI agents can safely learn, adapt, and improve. By combining high-fidelity simulation environments with state-of-the-art training loops, we build specialized models that solve long-horizon, complex tasks with unmatched reliability. We're building the most talent-dense AI team in Europe to make this happen.

You'll own the compute and execution platform supporting training, inference, task generation, and evaluation. Your work will make GPU resources productive and research workflows straightforward to operate and debug.

What you'll work on
  • Build and operate GPU clusters, job scheduling, networking, and storage for distributed ML workloads.
  • Develop the execution platform for large numbers of concurrent environments, including sandbox isolation, resource limits, retries, and state recovery.
  • Build reliable data and artifact pipelines for datasets, trajectories, model weights, and checkpoints.
  • Own platform observability and recovery; improve capacity allocation, startup latency, and reliability through measurable changes.
  • Give researchers reproducible environments and simple tools to launch, inspect, and debug experiments.
Relevant technologies
  • Compute and orchestration: Linux, Docker, Kubernetes, Slurm, and Terraform.
  • Execution and storage: Python or Go, Ray, S3-compatible object storage, and PostgreSQL.
  • Observability: Prometheus, Grafana, OpenTelemetry, and NVIDIA DCGM; familiarity with GPU networking and NCCL is mandatory.
About you

You have strong systems fundamentals and can reason carefully about concurrency, resource contention, and failure recovery. You enjoy making complex infrastructure understandable and dependable for the people using it.

We welcome infrastructure specialists and exceptionally fast-learning generalists. Prior ML infrastructure experience is valuable; evidence of building and operating reliable systems matters.

What we offer
  • Competitive compensation and equity.
  • A fast-paced environment combining frontier research with impactful real-world applications.
  • Visa sponsorship and relocation support for candidates joining us in Paris or London.
  • A flexible hybrid setup, with a preference for working together in person.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Member of Technical Staff — Platform Engineering
Member of Technical Staff — Platform Engineering

Remanence • Paris (TX)

Hybrid
USD 136,000 - 204,000
Visa sponsorship
Relocation support
Hybrid work setup
ML Infrastructure Engineer – GPU Compute Platform
ML Infrastructure Engineer – GPU Compute Platform

Remanence • Paris (TX)

Hybrid
USD 124,000 - 186,000
Visa sponsorship
Relocation support
Hybrid work setup
+1
Member of Technical Staff
Member of Technical Staff

Harrison Clarke • San Francisco (CA)

On-site
USD 180,000 - 280,000
Member of Technical Staff — Post-training / RL
Member of Technical Staff — Post-training / RL

Remanence • Paris (TX)

Hybrid
USD 135,000 - 203,000
Visa sponsorship
Relocation support
Hybrid setup
Member of Technical Staff — Compute Cluster
Member of Technical Staff — Compute Cluster

Causal • San Francisco (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff — Compute Cluster
Member of Technical Staff — Compute Cluster

Kindredventures • San Francisco (CA)

On-site
USD 140,000 - 230,000
Member of Technical Staff — Environments / Evals
Member of Technical Staff — Environments / Evals

Remanence • Paris (TX)

Hybrid
USD 102,000 - 147,000
Visa sponsorship
Relocation support
Hybrid work
Member of Technical Staff - Infrastructure
Member of Technical Staff - Infrastructure

Gimlet Labs • San Francisco (CA)

On-site
USD 120,000 - 160,000
Member of Technical Staff — Compute Cluster
Member of Technical Staff — Compute Cluster

Causal Labs • San Francisco (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff, Infrastructure
Member of Technical Staff, Infrastructure

Psi • Boston (MA), Northern (KY)

On-site
USD 180,000 - 260,000
Meaningful equity
Competitive compensation
Benefits