Remote Performance Engineer: ML Training & Kernels

Cohere

Greater London

On-site

GBP 75,000 - 95,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Co-working benefit
Daily lunch program
Regular community and social events

Job summary

Cohere is seeking a Performance Engineer in Greater London to optimize the performance of advanced language models and systems. You will focus on improving model training metrics and ensuring high accelerator utilization.

The role involves designing scalable software, writing low-level CUDA kernels, and experimenting with ideas on super-compute infrastructure. Applicants should have strong software engineering skills and proficiency in Python, with experience in GPUs and distributed training.

Cohere supports a remote-friendly work environment, offering benefits like co-working spaces for remote teams.

Qualifications

  • Extremely strong software engineering skills.
  • Proficiency in Python and related ML frameworks.
  • Experience writing kernels for GPUs using CUDA.

Responsibilities

  • Design and write high‑performant and scalable software for training.
  • Understand architectural modifications and their effects on training throughput.
  • Research and implement ideas on our super‑compute and data infrastructure.

Skills

Strong software engineering skills
Proficiency in Python
Experience with ML frameworks such as JAX
Experience writing kernels for GPUs using CUDA
Familiarity with autoregressive sequence models

Job description

Cohere is seeking a Performance Engineer in Greater London to optimize the performance of advanced language models and systems. You will focus on improving model training metrics and ensuring high accelerator utilization.

The role involves designing scalable software, writing low-level CUDA kernels, and experimenting with ideas on super-compute infrastructure. Applicants should have strong software engineering skills and proficiency in Python, with experience in GPUs and distributed training.

Cohere supports a remote-friendly work environment, offering benefits like co-working spaces for remote teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Performance Engineer: Scale GPU/CPU ML Workloads
ML Performance Engineer: Scale GPU/CPU ML Workloads

G-Research • Greater London

Hybrid
GBP 90,000 - 150,000
Competitive pay
Lunch provided
Annual leave 35d
+5
Senior ML Systems Engineer: Training Frameworks & HPC
Senior ML Systems Engineer: Training Frameworks & HPC

Visa Hunt • Greater London

Hybrid
GBP 110,000 - 170,000
Weekly lunch stipend
Health and dental benefits
RRSP matching / Pension
+5
ML Performance Engineer: Large-Scale GPU/CPU Optimization
ML Performance Engineer: Large-Scale GPU/CPU Optimization

gresearch • Greater London

On-site
GBP 90,000 - 130,000
Lunch provided
35 days annual leave
9% pension contributions
+3
Staff Engineer – Training Infra & ML Systems
Staff Engineer – Training Infra & ML Systems

Cohere • Greater London

On-site
GBP 70,000 - 90,000
Open and inclusive culture
Weekly lunch stipend of $75/£75
Full health and dental benefits
+2
Remote Staff ML Efficiency Engineer — Scale & Optimize
Remote Staff ML Efficiency Engineer — Scale & Optimize

Reddit, Inc. • Greater London

On-site
GBP 110,000 - 160,000
Global Benefits
Family Planning
Mental Health Support
+4
Embedded Performance Engineer: ML Kernel Optimisation
Embedded Performance Engineer: ML Kernel Optimisation

Fractile • Greater London

Hybrid
GBP 50,000 - 70,000
ML Performance Engineer – Scale GPU/CPU Workloads
ML Performance Engineer – Scale GPU/CPU Workloads

Barlowe LLP • Greater London

On-site
GBP 90,000 - 150,000
Lunch provided
35 days’ annual leave
9% company pension contributions
+4
GPU Systems Performance Engineer
GPU Systems Performance Engineer

Janestreet • Greater London

On-site
GBP 111,000 - 140,000
Senior ML Performance Engineer - Real-Time Inference & Scale
Senior ML Performance Engineer - Real-Time Inference & Scale

Odyssey • Greater London

On-site
GBP 70,000 - 90,000
ML Infrastructure Engineer - Distributed Systems & HPC
ML Infrastructure Engineer - Distributed Systems & HPC

Stanford Black Limited • Greater London

On-site
GBP 100,000 - 150,000
Competitive compensation
Bonus structure
Autonomy from day one
+1