LLM Performance Engineer — Scale, Optimize & Deploy

Isomorphic Labs

Greater London

Hybrid

GBP 120,000 - 160,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Isomorphic Labs in the United Kingdom is seeking an experienced ML engineer to scale frontier models in production. You will join the model performance and scaling team, partnering with scientists and engineers to accelerate model training and deployment for drug discovery.

You will apply novel techniques to performance optimization, leverage large-scale distributed training, and translate research advances into robust, production‑ready systems that impact real-world biopharma projects.

Qualifications

  • Significant experience with large scale distributed training of LLMs.
  • Experience with deep learning ML frameworks (either JAX or PyTorch).
  • Knowledge of parallelism strategies and collective communication libraries (e.g. NCCL).
  • Good understanding of GPU architectures. Reasoning about performance concepts is more important than writing kernels from scratch
  • Excellent collaboration skills.

Responsibilities

  • Implement and optimize LLM post-training methods at scale on frontier models.
  • Collaborate with research teams to translate new methods into production-ready systems.
  • Relentlessly prioritize and execute on performance optimization opportunities.
  • Evaluate and deploy frameworks for supervised fine-tuning, reinforcement learning and LLM evaluation.
  • Diagnose and fix performance bottlenecks and communication overhead in distributed training and inference systems.
  • Deploy low-precision methods to balance performance with accuracy, impacting real world drug design programs.

Skills

Distributed training of LLMs
DL frameworks (JAX or PyTorch)
Parallelism & NCCL
GPU architectures
Collaboration

Job description

Isomorphic Labs in the United Kingdom is seeking an experienced ML engineer to scale frontier models in production. You will join the model performance and scaling team, partnering with scientists and engineers to accelerate model training and deployment for drug discovery.

You will apply novel techniques to performance optimization, leverage large-scale distributed training, and translate research advances into robust, production‑ready systems that impact real-world biopharma projects.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLM Performance Engineer — Scale Frontiers in AI
LLM Performance Engineer — Scale Frontiers in AI

Isomorphic Labs • Greater London

Hybrid
GBP 80,000 - 120,000
Frontier AI LLM Systems Engineer - London
Frontier AI LLM Systems Engineer - London

isomorphiclabs • Greater London

On-site
GBP 90,000 - 150,000
Remote MLOps Engineer for Scalable Molecular Modeling
Remote MLOps Engineer for Scalable Molecular Modeling

Boltz • Greater London

Hybrid
GBP 90,000 - 140,000
High-Throughput ML Systems Performance Engineer
High-Throughput ML Systems Performance Engineer

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 150,000
Health insurance
Fertility benefits
Parental leave 22 weeks
+12
ML Platform Lead - LLM Training & Inference
ML Platform Lead - LLM Training & Inference

Scale AI • York and North Yorkshire

On-site
GBP 120,000 - 180,000
Health & Wellbeing
Career Growth stipend
Community events
+1
Research Engineer (LLM Performance), London New London
Research Engineer (LLM Performance), London New London

Isomorphic Labs • Greater London

On-site
GBP 80,000 - 120,000
ML Data & Platform Engineer — Scale Training Pipelines
ML Data & Platform Engineer — Scale Training Pipelines

Speechmatics • England

Hybrid
GBP 75,000 - 110,000
Private Medical
Dental
Global opportunities
+4
ML Performance Engineer – Scale GPU/CPU Workloads
ML Performance Engineer – Scale GPU/CPU Workloads

Barlowe LLP • Greater London

On-site
GBP 90,000 - 150,000
Lunch provided
35 days’ annual leave
9% company pension contributions
+4
NLP Performance Engineer: Scale LLM Inference for Research
NLP Performance Engineer: Scale LLM Inference for Research

Barlowe LLP • Greater London

On-site
GBP 90,000 - 150,000
Competitive compensation + bonus
Lunch provided
35 days annual leave
+5
Senior LLM Serving Platform Engineer
Senior LLM Serving Platform Engineer

Scale AI • Greater London

On-site
GBP 90,000 - 130,000