Performance RL Research Engineer - Safe Code & AI (Hybrid)
Anthropic
San Francisco (CA)
Hybrid
USD 350,000 - 850,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
Competitive compensation
Generous vacation and parental leave
Flexible working hours
Office space for collaboration
Job summary
A leading AI research company in San Francisco is seeking a Research Engineer for their Code RL team. The successful candidate will advance AI models' ability to write correct and efficient code while working in collaboration with teams across the organization. Key requirements include expertise with accelerators, ML frameworks, and a bachelor's degree or equivalent experience. The role offers a competitive salary ranging from $350,000 to $850,000, with a hybrid work model.
Qualifications
Expertise with accelerators and ML frameworks needed.
Experience across the stack from kernels to model code is ideal.
Commitment to developing safe AI systems is required.
Responsibilities
Invent, design, and implement RL environments and evaluations.
Conduct experiments and shape the research roadmap.
Deliver work into training runs and collaborate cross-functionally.
Skills
Expertise with accelerators (CUDA, ROCm, Triton, Pallas)
ML framework programming (JAX or PyTorch)
Balancing research exploration with engineering implementation
Passion for AI's potential
Education
Bachelor's degree in a related field or equivalent experience
Tools
CUDA
ROCm
Triton
Pallas
JAX
PyTorch
Job description
A leading AI research company in San Francisco is seeking a Research Engineer for their Code RL team. The successful candidate will advance AI models' ability to write correct and efficient code while working in collaboration with teams across the organization. Key requirements include expertise with accelerators, ML frameworks, and a bachelor's degree or equivalent experience. The role offers a competitive salary ranging from $350,000 to $850,000, with a hybrid work model.