Research Engineer - Post-Training

Voltai

Palo Alto (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Voltai is seeking engineers to post-train frontier models that autonomously handle complex semiconductor tasks. You will design RL environments, generate rich evaluation datasets, and partner with hardware and verification experts to define metrics and execution conditions.

You will craft reward functions, scaling strategies, and rigorous evaluation frameworks to push models toward reliable, efficient, and creative silicon reasoning in next-generation designs.

Qualifications

  • Experience creating and scaling RL environments for LLMs or multimodal agents.
  • Ability building evaluation datasets and benchmarks for complex reasoning or design tasks.
  • Collaborating with hardware/verification experts to define evaluation metrics and conditions.
  • Experience designing reward functions and feedback pipelines balancing correctness and efficiency.
  • Conducting large-scale RL fine-tuning or post-training experiments for frontier models.

Responsibilities

  • Post-train frontier models to autonomously perform complex tasks in semiconductor design and verification.
  • Develop RL environments that simulate chip design workflows and verification steps.
  • Create evaluation metrics, scaling strategies, and frameworks to improve model reliability and creativity in silicon reasoning.

Skills

RL environments
Evaluation datasets
Collaboration with hardware
Reward design
RL fine-tuning
Curriculum learning

Job description

About Voltai

Voltai is developing world models, and agents to learn, evaluate, plan, experiment, and interact with the physical world. We are starting out with understanding and building hardware; electronics systems and semiconductors where AI can design and create beyond human cognitive limits.

About the Team

Backed by Silicon Valley’s top investors, Stanford University, and CEOs/Presidents of Google, AMD, Broadcom, Marvell, etc. We are a team of previous Stanford professors, SAIL researchers, Olympiad medalists (IPhO, IOI, etc.), CTOs of Synopsys & GlobalFoundries, Head of Sales & CRO of Cadence, former US Secretary of Defense, National Security Advisor, and Senior Foreign-Policy Advisor to four US presidents.

Post-Training

In this role, you will post-train frontier models to autonomously perform complex tasks across the semiconductor design and verification pipeline. Models you train will propose and optimize chip architectures, generate and refine RTL code, run simulations, identify verification gaps, and iteratively improve designs — accelerating the pace of semiconductor innovation.

You will collaborate with leading experts in hardware design, verification, and computer architecture to design rich reinforcement learning environments that capture the intricacies of chip design workflows. You’ll develop structured reward functions, scaling strategies, and evaluation frameworks that push models toward higher reliability, efficiency, and creativity in semiconductor reasoning.

Your work will directly advance the goal of creating AI systems capable of reasoning about, designing, and verifying next-generation silicon systems.

You might thrive in this role if you have experience with
  • Creating and scaling RL environments for LLMs or multimodal agents
  • Building high-quality evaluation datasets and benchmarks for complex reasoning or design tasks
  • Working closely with domain experts in hardware and verification to define evaluation metrics, constraints, and simulation conditions
  • Designing reward functions and feedback pipelines that balance correctness, performance, and design efficiency
  • Running large-scale RL fine-tuning or post-training experiments for frontier models
  • Applying reinforcement learning or curriculum learning to structured reasoning or symbolic domains
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Research Engineer - Mid-Training
Research Engineer - Mid-Training

Voltai • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Research Engineer: AI for Chip Design & RL
Research Engineer: AI for Chip Design & RL

Voltai • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Machine Learning Engineer
Machine Learning Engineer

Voltai • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Design Verification Engineer
Design Verification Engineer

Voltai • Palo Alto (CA)

On-site
USD 150,000 - 210,000
ML Engineer – Multimodal AI & LLM Production
ML Engineer – Multimodal AI & LLM Production

Voltai • Palo Alto (CA)

On-site
USD 180,000 - 240,000
AI Research Engineer - Semiconductor Design & Verification
AI Research Engineer - Semiconductor Design & Verification

Voltai • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Software Engineer
Software Engineer

Voltai • United States

On-site
USD 120,000 - 180,000
Research Engineer - CUDA Kernel Engineering
Research Engineer - CUDA Kernel Engineering

Voltai • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Formal Verification Research Scientist
Formal Verification Research Scientist

Voltai • Palo Alto (CA)

On-site
USD 140,000 - 210,000
Hardware Application Engineer
Hardware Application Engineer

Voltai • Palo Alto (CA)

On-site
USD 150,000 - 210,000