Research Scientist (post-training)

Genmo

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading AI research lab is seeking a Research Scientist to focus on alignment and post-training techniques for advanced video generation models. The role involves designing evaluation frameworks, collaborating with teams, and mentoring researchers to enhance alignment methods. Candidates should possess a Ph.D. and extensive experience in reinforcement learning and large-scale training systems. Located in San Francisco, the position offers a chance to be at the forefront of generative AI innovation.

Qualifications

  • Ph.D. in Computer Science, Artificial Intelligence, Machine Learning, or closely related field.
  • Strong publication record in top-tier conferences.
  • Extensive experience in optimizing large-scale training pipelines.

Responsibilities

  • Lead research initiatives in alignment and post-training methods for models.
  • Design and implement supervised fine-tuning and RLHF pipelines.
  • Collaborate with cross-functional teams on alignment improvements.

Skills

Reinforcement learning
Aligning generative models
Implementation with PyTorch
Collaborative skills
Evaluation frameworks design

Education

Ph.D. in relevant field

Tools

PyTorch
Distributed training systems

Job description

Overview

We are Genmo, a research lab dedicated to building open, state-of-the-art models for video generation towards unlocking the right brain of AGI. Join us in shaping the future of AI and pushing the boundaries of what's possible in video generation.

Role overview

We are seeking an exceptional Research Scientist to join our team, focusing on alignment and post-training techniques for large-scale video generation models. In this role, you will be at the forefront of ensuring our diffusion-based video models reliably produce high-quality, physically accurate and safe outputs that match human preferences and values.

Responsibilities
  • Lead research initiatives in alignment and post-training methods for video generation models, focusing on improved quality, reliability, and adherence to human intent

  • Design and implement supervised fine-tuning and reinforcement learning from human feedback (RLHF) pipelines for video generation models

  • Develop robust evaluation frameworks to measure model alignment, safety, and output quality

  • Create and optimize data collection pipelines for human feedback and preferences

  • Design and conduct experiments to validate alignment techniques and their scaling properties

  • Collaborate with cross-functional teams to integrate alignment improvements into our production pipeline

  • Stay at the cutting edge of the field by regularly reviewing academic literature in both generative AI and alignment

  • Mentor junior researchers and foster a culture of responsible AI development

  • Work closely with product teams to ensure alignment methods enhance rather than inhibit model capabilities

Qualifications
  • Ph.D. in Computer Science, Artificial Intelligence, Machine Learning, or a closely related field

  • Must have:

    • Strong publication record in top-tier conferences (e.g., NeurIPS, ICML, ICLR) with a focus on reinforcement learning, alignment, or generative models

    • Extensive experience implementing and optimizing large-scale training pipelines using PyTorch

    • Deep understanding of reinforcement learning techniques, particularly RLHF

    • Experience with distributed training systems and large-scale experiments

    • Proven track record in designing and implementing robust evaluation frameworks

    • Excellent communication skills with the ability to explain complex technical concepts to diverse audiences

    • Strong software engineering skills and experience with complex shared codebases

  • Ideal candidate will have:

    • Experience with diffusion models or other generative architectures

    • Background in fine-tuning large language models or generative models

    • Experience working with human feedback data collection and annotation pipelines

    • Strong aesthetic sense and understanding of video quality assessment

    • Familiarity with alignment techniques such as constitutional AI or debate

    • Track record of successful collaboration with product teams

    • Experience with perceptual quality metrics and human evaluation design

    • Contributions to open-source projects in AI alignment or generative AI

    • Additional Information

    Additional Information

    The role is based in the Bay Area (San Francisco). Candidates are expected to be located near the Bay Area or open to relocation.

    Genmo is an Equal Opportunity Employer. Candidates are evaluated without regard to age, race, color, religion, sex, disability, national origin, sexual orientation, veteran status, or any other characteristic protected by federal or state law. Genmo, Inc. is an E-Verify company and you may review the Notice of E-Verify Participation and the Right to Work posters in English and Spanish.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Scientist (diffusion)
Research Scientist (diffusion)

Genmo • San Francisco (CA)

On-site
USD 120,000 - 160,000
Software Engineer (New Grad)
Software Engineer (New Grad)

algojobs • San Francisco (CA)

On-site
USD 110,000 - 150,000
Fast-paced startup environment
Opportunity for growth and learning
Collaborative team culture
Research Scientist - Multimodal Agent, Consumer Devices
Research Scientist - Multimodal Agent, Consumer Devices

OpenAI • United States

Hybrid
USD 140,000 - 230,000
Relocation assistance
Hybrid work model (3 days in office)
Research Engineer/Scientist - Human Alignment, Consumer Devices
Research Engineer/Scientist - Human Alignment, Consumer Devices

SupportFinity™ • San Francisco (CA)

On-site
USD 100,000 - 180,000
Research Engineer/Scientist - Human Alignment, Consumer Devices
Research Engineer/Scientist - Human Alignment, Consumer Devices

OpenAI • San Francisco (CA)

Hybrid
USD 380,000 - 445,000
Research Scientist (Generative Modeling)
Research Scientist (Generative Modeling)

WORLD LABS • San Francisco (CA)

On-site
USD 250,000 - 325,000
Research, Post-Training Data
Research, Post-Training Data

Thinkingmachines • San Francisco (CA)

On-site
USD 350,000 - 475,000
Health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
Research Scientist, Multimodal Alignment, Safety, and Fairness
Research Scientist, Multimodal Alignment, Safety, and Fairness

Google DeepMind • Town of Kirkland (NY)

On-site
USD 147,000 - 211,000
Competitive salary
Bonuses
Equity options
+1
Research Scientist, World Models Graduate (Intelligent Creation) - Global Frontier Tech Recruit[...]
Research Scientist, World Models Graduate (Intelligent Creation) - Global Frontier Tech Recruit[...]

TikTok • San Jose (CA)

On-site
USD 244,800 - 588,000
Medical, dental, and vision insurance
401(k) savings plan with company match
Paid parental leave
+5
Video Generation AI Research Scientist — Alignment & RLHF
Video Generation AI Research Scientist — Alignment & RLHF

Genmo • San Francisco (CA)

On-site
USD 120,000 - 160,000