Software Engineer, RL Environments

Cursor

San Francisco (CA)

On-site

USD 150,000 - 230,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

SpaceXAI is seeking a Software Engineer to join the RL Environments team. You will build platforms that enable scalable, realistic reinforcement learning environments and tasks, working across data, research, and engineering to ship trustworthy data quickly.

You will design end-to-end data factories, enforce quality standards, and develop self-serve APIs and tooling for internal teams and external contributors, partnering across disciplines to translate model-capability goals into training tasks.

Qualifications

  • Demonstrated strong software engineering fundamentals.
  • Experience with data platforms, developer tools, distributed systems, or ML infrastructure.
  • Ability to turn ambiguous quality standards into automated checks.
  • Ability to build repeatable pipelines from messy data.
  • Collaborates across disciplines end-to-end.

Responsibilities

  • Build platforms to create realistic RL environments at scale.
  • Design end-to-end factories turning raw data into environments.
  • Define and apply quality standards across data sources.
  • Create self-serve APIs and tooling for internal/external users.
  • Collaborate with research, platform, and external teams to translate goals into tasks and environments.

Skills

Software engineering fundamentals
Data platforms
Developer tools
Distributed systems
ML infrastructure
Quality automation checks
Cross-disciplinary collaboration

Tools

APIs
Self-serve tooling

Job description

Our mission is to automate coding. The first step in our journey is to build the best tool for professional programmers, using a combination of inventive research, design, and engineering. Our organization is very flat, and our team is small and talent dense. We particularly like people who are truth-seeking, passionate, and creative. We enjoy spirited debate, crazy ideas, and shipping code.

About The Role

One of our north stars is a future where teams can hire Grok as a remote colleague. Reaching that future requires training our models in realistic, diverse environments that support complex, end-to-end work across domains.

As a Software Engineer on the RL Environments team at SpaceXAI, you’ll build the systems that turn real-world data into high-quality environments and tasks for reinforcement learning. You’ll create dramatically more realistic training environments while owning the shared platforms and quality layers that help us produce trustworthy data quickly and at scale.

This role sits at the intersection of research, data, and engineering. You’ll work closely with ML Platform and research teams to turn company data, Grok Bot interactions, vendor-built tasks, tutor data, acquired data, and synthetic data into training-ready environments. Your work will make high-quality RL data easier to create, validate, discover, and use across our model-training efforts.

What you’ll work on
  • Building platforms that allows us to build complex, realistic, and diverse RL environments at scale, that support end-to-end tasks across knowledge-work domains.
  • Designing end-to-end factory that turns massive raw data into useful and realistic environments.
  • Defining and applying consistent quality standards across vendor, tutor, acquired, and synthetic data.
  • Creating self-serve APIs and tooling that accelerate task and environment development for internal teams and external contributors.
  • Partnering across research, platform, and external teams to translate model-capability goals into effective training tasks and environments.
You may be a fit if
  • You have strong software engineering fundamentals and experience with data platforms, developer tools, distributed systems, or ML infrastructure.
  • You can turn ambiguous quality standards into concrete, automated checks.
  • You’re comfortable building repeatable pipelines from messy, heterogeneous data.
  • You care about model behavior and can translate capability goals into tasks and experiments.
  • You collaborate well across disciplines and own open-ended problems end to end.
Applying

If there appears to be a fit, we'll reach out to schedule 2-3 short technicals. After, we'll schedule an onsite in our office, where you'll work on a small project, discuss ideas, and meet the team.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

RL Environments Engineer: Realistic Training Tasks
RL Environments Engineer: Realistic Training Tasks

Cursor • San Francisco (CA)

On-site
USD 150,000 - 230,000
RL Environment Software Engineer
RL Environment Software Engineer

talentpluto • San Francisco (CA)

On-site
USD 180,000 - 220,000
Software Engineer, Research Tools
Software Engineer, Research Tools

Cursor • San Francisco (CA)

On-site
USD 150,000 - 190,000
RL Environments Engineer
RL Environments Engineer

Bespoke Labs Inc. • Mountain View (CA), Northern (KY)

Hybrid
USD 250,000 - 300,000
Health, dental, and vision coverage
401(k)
Daily onsite lunch provided
+2
RL Environments Engineer
RL Environments Engineer

Bespoke-Labs • Mountain View (CA)

On-site
USD 250,000 - 300,000
Health, dental, and vision
401(k)
Daily onsite lunch
+3
Full Stack Software Engineer - RL environments
Full Stack Software Engineer - RL environments

Acceler8 Talent • San Francisco (CA)

On-site
USD 120,000 - 180,000
Backend Engineer
Backend Engineer

bespokelabs • Mountain View (CA)

On-site
USD 120,000 - 160,000
Health coverage
Competitive salary and equity
Opportunity to work with leading AI research labs
Staff Software Engineer, RL Environments
Staff Software Engineer, RL Environments

Scale AI, Inc. • New York (NY)

On-site
USD 252,000 - 315,000
Equity
Health benefits
Learning stipend
+1
Staff Software Engineer, RL Environments
Staff Software Engineer, RL Environments

Scale AI, Inc. • San Francisco (CA)

On-site
USD 252,000 - 315,000
Health coverage
Dental coverage
Vision coverage
+4
Reinforcement Learning Environment Engineer
Reinforcement Learning Environment Engineer

Open Data Science • San Francisco (CA)

Remote
USD 100,000 - 150,000