Staff AI Security Engineer - Post-Training & Scale

Armadin

Palo Alto (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Full health, dental, vision
Equity ownership
In-office meals
Haircuts at the office
Conferences & events
401(k), HSA, and FSA plans
Flexible PTO

Job summary

Armadin, based in Palo Alto, is seeking a senior researcher to own the post-training strategy and to translate security research into scalable model capabilities. You will drive fine-tuning, reward modeling, and efficiency benchmarks, ensuring reliability and impact across production environments.

The role requires a PhD or equivalent experience with hands-on SFT, RLHF/RLAIF, and distillation. You will collaborate across teams to push frontier research into practical, enterprise-grade solutions.

Qualifications

  • PhD in CS, ML, or related field, or equivalent experience with a strong track record.
  • Hands-on experience with SFT, RLHF/RLAIF, DPO, and reward modeling.
  • Practical ability to train, fine-tune, and evaluate models and deploy research to production.
  • Engineering strength to implement ideas and run them at scale.
  • Ability to drive strategy broadly or focus deeply on hard problems as needed.
  • Strong collaboration across teams and focus on outcomes.

Responsibilities

  • Own Post-Training Strategy: Drive the post-training roadmap (fine-tuning, preference optimization, reward modeling, RL, distillation) to make models more capable, reliable, and aligned.
  • Push Efficiency & Evals: Make models serve reliably at scale, and build the benchmarks that measure quality and catch regressions.
  • Ground Research in Reality: Turn our security experts’ tradecraft into model capabilities and ship your techniques into production.
  • Stay at the Frontier: Track post-training and efficiency research and bring the best ideas in.

Skills

PhD or equivalent
Post-training techniques
Model training & evaluation
Scale engineering
Strategic thinking
Cross-team collaboration

Education

PhD in CS/ML or related field

Job description

Armadin, based in Palo Alto, is seeking a senior researcher to own the post-training strategy and to translate security research into scalable model capabilities. You will drive fine-tuning, reward modeling, and efficiency benchmarks, ensuring reliability and impact across production environments.

The role requires a PhD or equivalent experience with hands-on SFT, RLHF/RLAIF, and distillation. You will collaborate across teams to push frontier research into practical, enterprise-grade solutions.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Member of Technical Staff, AI Security & Agents
Senior Member of Technical Staff, AI Security & Agents

Armadin • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Full Health, Dental, & Vision Coverage
Meaningful Equity Ownership
In-Office Meals
+4
Senior Technical Recruiter — AI/ML Talent for Security
Senior Technical Recruiter — AI/ML Talent for Security

Armadin • Palo Alto (CA)

On-site
USD 140,000 - 190,000
Health coverage
Equity ownership
In-office meals
+4
AI Security Research Engineer: RL Training & Red Teaming
AI Security Research Engineer: RL Training & Red Teaming

Qualis • Sunnyvale (CA)

On-site
USD 120,000 - 180,000
Member of Technical Staff - Research
Member of Technical Staff - Research

Armadin • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Full health, dental, vision
Equity ownership
In-office meals
+4
Software Engineer Intern — Autonomous Security Platform
Software Engineer Intern — Autonomous Security Platform

Armadin • Palo Alto (CA)

On-site
USD 88,000 - 118,000
Health, dental, and vision coverage
Equity ownership
In-office meals
+4
AI Research Engineer
AI Research Engineer

Qualis • Sunnyvale (CA)

On-site
USD 120,000 - 180,000
Senior AI Security Engineer: Red Team & Secure Inference
Senior AI Security Engineer: Red Team & Secure Inference

Jobtailor • Pennsylvania

On-site
USD 180,000 - 260,000
Staff AI Safety Engineer — Red Team & Guardrails
Staff AI Safety Engineer — Red Team & Guardrails

Visa Hunt • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 260,000
Top-tier compensation
Stock options
Health & wellness
+3
Senior Red Team Operator & Offensive Security Leader
Senior Red Team Operator & Offensive Security Leader

Armadin • Palo Alto (CA)

On-site
USD 200,000 - 320,000
Full Health, Dental, & Vision Coverage
Meaningful Equity Ownership
In-Office Meals
+3
AI Security Researcher: LLM Post-Training & Evaluations
AI Security Researcher: LLM Post-Training & Evaluations

Palo Alto Networks • California (MO)

Hybrid
USD 163,000 - 264,000