ML Scientist - Adversarial Robustness

Mercor

San Francisco (CA)

On-site

USD 180,000 - 235,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Mercor is seeking experienced machine learning researchers to advance deep learning models across vision and language, training end-to-end and improving empirical open-ended ML research outcomes. You will take ownership of scalable training, robustness, and model compression while addressing hard constraints on data, compute, and latency.

The role emphasizes hands-on experimentation with adversarial robustness, efficient computer vision, and multilingual pre-training, alongside LLM post-training

Qualifications

  • 3+ years of machine learning research experience (PhD research counts toward this requirement).
  • Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks.
  • Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions.

Responsibilities

  • Train image classifiers and generative image models from scratch, and fine-tune open-weight language models.
  • Get the most out of limited data, compute, and model-size budgets.
  • Make models robust — to adversarial inputs and to adversarial conversations.
  • Compress models to meet hard size and latency constraints without sacrificing accuracy.
  • Diagnose and resolve training issues.

Skills

Adversarial Robustness
Efficient Computer Vision
Generative Image Modeling
LLM Post-Training & Behavioral Robust
Multilingual Pre-training

Education

PhD research counts toward requirement

Tools

PyTorch
JAX
TensorFlow

Job description

We're looking for experienced machine learning researchers with hands-on experience training and improving deep learning models end-to-end, across vision and language. You'll work on well-scoped empirical open-ended ML research problems.

Responsibilities
  • Train image classifiers and generative image models from scratch, and fine-tune open-weight language models.
  • Get the most out of limited data, compute, and model-size budgets.
  • Make models robust — to adversarial inputs and to adversarial conversations.
  • Compress models to meet hard size and latency constraints without sacrificing accuracy.
  • Diagnose and resolve training issues.
Requirements

We are looking for candidates with strong expertise in one or more of the following areas:

Adversarial Robustness

Experience with:

  • Adversarial training of image classifiers (e.g. PGD-based training, TRADES).
  • Evaluating robust accuracy under standard threat models (e.g. L∞ attacks, AutoAttack) and avoiding gradient-masking pitfalls.
  • Managing the robustness–accuracy trade-off and robust overfitting.
Efficient Computer Vision

Experience with:

  • Training image classifiers end-to-end, especially for fine-grained recognition (many visually similar classes, few examples per class).
  • Model compression: quantization, pruning, and knowledge distillation from large teachers into small students.
  • Deploying models under hard size or latency budgets (on-device, edge, or embedded settings).
Generative Image Modeling

Experience with:

  • Training image generative models from scratch: diffusion models, GANs, VAEs, or flow-based models.
  • Iterating against sample-quality metrics such as FID.
  • Training-efficiency tricks that produce good generators quickly and at small parameter counts.
LLM Post-Training & Behavioral Robustness

Hands-on experience with one or more of:

  • Supervised fine-tuning and preference optimisation (DPO, RLHF, RLAIF) of open-weight language models, including building your own datasets via synthetic generation, noisy or weak supervision, and rejection sampling.
  • Shaping conversational behaviour over multiple turns: resistance to persuasion and sycophancy, calibrated confidence, and knowing when to accept corrections.
  • Alignment-style fine-tuning that changes a specific behaviour while preserving general capability.
Multilingual Pre-training

Experience with:

  • Training multilingual or low-resource-language models from scratch.
  • Tokenizer design across scripts and typologically diverse languages.
  • Balancing highly unequal per-language data (sampling temperatures, cross-lingual transfer) in data-constrained regimes.
Additional Areas of Interest

Experience in any of the following is a plus:

  • Scaling laws and training-efficiency research.
  • Curriculum learning and data ordering.
  • Model evaluation: benchmark construction, contamination control, statistically sound comparisons.
  • Uncertainty estimation and model calibration.
  • Data augmentation and synthetic data for robustness.
General Qualifications
  • 3+ years of machine learning research experience (PhD research counts toward this requirement).
  • Strong experience with PyTorch, JAX, TensorFlow, or similar ML frameworks.
  • Degree from a top-100 university, experience at a FAANG or comparable AI company, or an equivalent research track record through publications or impactful open-source contributions.
Why Join
  • Work on cutting-edge machine learning research.
  • Collaborate with leading AI researchers on challenging, high-impact projects.
  • Flexible, project-based work with competitive compensation.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Scientist - Adversarial Robustness - AI Trainer
ML Scientist - Adversarial Robustness - AI Trainer

Mercor • Chicago (IL)

On-site
USD 120,000 - 180,000
ML Scientist - Adversarial Robustness - AI Trainer
ML Scientist - Adversarial Robustness - AI Trainer

Obsidian • Chicago (IL)

On-site
USD 130,000 - 190,000
ML Scientist - Adversarial Robustness
ML Scientist - Adversarial Robustness

Obsidian • San Francisco (CA)

On-site
USD 180,000 - 260,000
Flexible, project-based work
Competitive compensation
ML Scientist - Adversarial Robustness - AI Trainer
ML Scientist - Adversarial Robustness - AI Trainer

Obsidian • San Diego (CA)

On-site
USD 150,000 - 210,000
ML Scientist - Adversarial Robustness - AI Trainer
ML Scientist - Adversarial Robustness - AI Trainer

Mercor • San Diego (CA)

On-site
USD 180,000 - 240,000
Adversarial ML Scientist: Robustness & Efficient Models
Adversarial ML Scientist: Robustness & Efficient Models

Obsidian • San Francisco (CA)

On-site
USD 180,000 - 260,000
Flexible, project-based work
Competitive compensation
Adversarial Robustness ML Scientist – Efficient DL
Adversarial Robustness ML Scientist – Efficient DL

Mercor • San Diego (CA)

On-site
USD 180,000 - 240,000
Adversarial ML Scientist: Robustness & Efficiency
Adversarial ML Scientist: Robustness & Efficiency

Mercor • San Francisco (CA)

On-site
USD 180,000 - 235,000
ML Scientist: Adversarial Robustness & Efficient Vision
ML Scientist: Adversarial Robustness & Efficient Vision

Obsidian • Chicago (IL)

On-site
USD 130,000 - 190,000
Machine Learning Researcher – LLM
Machine Learning Researcher – LLM

Susquehanna International Group • Bala Cynwyd (PA)

On-site
USD 180,000 - 280,000