Remote Reinforcement Learning Systems Engineer

Bugcrowd

United States

On-site

USD 146,000 - 243,000

Full time

7 days ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Bugcrowd's RL and Reasoning Team builds autonomous cybersecurity environments to train next-gen AI systems. As a Staff Engineer, you will advance reinforcement learning development and deliver the infrastructure and tooling that turn real-world vulnerability research into scalable training environments.

This role focuses on pipelines that ingest software projects, analyze them with Bugcrowd’s Mayhem platform, and automatically construct thousands of RL environments used by frontier AI labs.

Qualifications

  • Understanding of RL training workflows used by modern LLM systems.
  • Experience with DevOps pipelines (e.g., github actions), reproducible builds (docker, buildkit, nix).
  • Proficiency in Python and C. Other languages (especially Rust) are a plus.
  • Understanding of software vulnerabilities, fuzzing, or program analysis
  • Experience with build systems and large open-source codebases
  • Comfort working with Linux systems and low-level debugging
  • Experience working with benchmark environments (CTFs, SWE-bench, security challenges, etc.)

Responsibilities

  • If you enjoy building high-performance systems that power cutting-edge AI research, this role is for you.
  • This role focuses on building the systems that generate RL environments, not just the environments themselves. You will design pipelines that ingest software projects, analyze them with Bugcrowd’s Mayhem platform, and automatically construct training environments used by frontier AI labs including Anthropic, OpenAI, and Cohere.
  • The ideal candidate is a strong systems engineer who understands Reinforcement learning workflows, system security, and low-level debugging.
  • You will build infrastructure that generates thousands of training environments used to train frontier AI systems.

Skills

RL workflows
Security background
Python
C
Rust

Tools

Mayhem platform

Job description

Bugcrowd's RL and Reasoning Team builds autonomous cybersecurity environments to train next-gen AI systems. As a Staff Engineer, you will advance reinforcement learning development and deliver the infrastructure and tooling that turn real-world vulnerability research into scalable training environments.

This role focuses on pipelines that ingest software projects, analyze them with Bugcrowd’s Mayhem platform, and automatically construct thousands of RL environments used by frontier AI labs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote RL Engineer for Cybersecurity AI Training
Remote RL Engineer for Cybersecurity AI Training

Bugcrowd • United States

On-site
USD 176,000 - 243,000
Senior AI RL Engineer — Systems & Infrastructure
Senior AI RL Engineer — Systems & Infrastructure

Mechanize • Oakland (CA), San Francisco (CA)

On-site
USD 180,000 - 240,000
Health insurance
Dental insurance
Vision insurance
+1
Junior Reinforcement Learning Engineer
Junior Reinforcement Learning Engineer

Mechanize • Oakland (CA), San Francisco (CA)

On-site
USD 90,000 - 150,000
Health insurance
Dental insurance
Vision insurance
+1
Remote Reinforcement Learning Engineer — Scale & Deploy
Remote Reinforcement Learning Engineer — Scale & Deploy

Bright Vision Technologies • Reston (VA)

On-site
USD 100,000 - 150,000
Remote Senior AI Software Engineer - RL Environments
Remote Senior AI Software Engineer - RL Environments

YO AI Labs • Town of Texas (WI)

Remote
USD 83,000 - 152,000
Remote Senior Software Engineer — RL Environment Designer
Remote Senior Software Engineer — RL Environment Designer

YO AI Labs • Los Angeles (CA)

Remote
USD 40,000 - 70,000
Remote Senior AI Training Engineer - RL Environments
Remote Senior AI Training Engineer - RL Environments

YO AI Labs • Philadelphia

Remote
USD 83,000 - 124,000
Remote work
Remote Reinforcement Learning Engineer — Scale & Deploy
Remote Reinforcement Learning Engineer — Scale & Deploy

Bright-Vision-Technologies • United States

Remote
USD 96,000 - 120,000
Senior Software Engineer - RL Environments (Remote)
Senior Software Engineer - RL Environments (Remote)

YO AI Labs • Boston (MA)

Remote
USD 83,000 - 138,000
Remote work
Senior Software Engineer (Remote) — RL Environments for AI
Senior Software Engineer (Remote) — RL Environments for AI

YO AI Labs • San Francisco (CA)

Remote
USD 83,000 - 124,000