Staff Performance Engineer - ML Training & GPU Kernel Expert

Deepstreamtech

Greater London

On-site

GBP 70,000 - 90,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Deepstreamtech is seeking a Performance Engineer in Greater London to enhance the performance of their advanced language models. The role demands strong software engineering and machine learning expertise, focusing on optimizing training metrics and developing efficient systems. You will work on CUDA and other tools to drive innovation in natural language processing and collaborate with leading researchers in the field.

Qualifications

  • Extremely strong software engineering skills.
  • Proficiency in Python and related ML frameworks such as JAX, Pytorch.
  • Experience writing kernels for GPUs using CUDA, triton.

Responsibilities

  • Optimize performance of advanced language models and systems.
  • Improve key model training metrics like training throughput.
  • Design high-performant software for training.

Skills

Software engineering skills
Proficiency in Python
Experience with ML frameworks (JAX, Pytorch)
Experience with GPU programming
Experience with distributed training
Familiarity with Transformers
Research paper publication

Job description

Deepstreamtech is seeking a Performance Engineer in Greater London to enhance the performance of their advanced language models. The role demands strong software engineering and machine learning expertise, focusing on optimizing training metrics and developing efficient systems. You will work on CUDA and other tools to drive innovation in natural language processing and collaborate with leading researchers in the field.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Remote Performance Engineer: ML Training & Kernels
Remote Performance Engineer: ML Training & Kernels

Cohere • Greater London

On-site
GBP 75,000 - 95,000
Co-working benefit
Daily lunch program
Regular community and social events
Senior ML Performance Engineer - Real-Time Inference & Scale
Senior ML Performance Engineer - Real-Time Inference & Scale

Odyssey • Greater London

On-site
GBP 70,000 - 90,000
Performance Engineer (GPU)
Performance Engineer (GPU)

Anthropic • York and North Yorkshire

On-site
GBP 90,000 - 140,000
Comprehensive health insurance
Fertility benefits
22 weeks parental leave
+1
Embedded Performance Engineer: ML Kernel Optimisation
Embedded Performance Engineer: ML Kernel Optimisation

Fractile • Greater London

Hybrid
GBP 50,000 - 70,000
High-Throughput ML Systems Performance Engineer
High-Throughput ML Systems Performance Engineer

Anthropic • York and North Yorkshire

On-site
GBP 110,000 - 150,000
Health insurance
Fertility benefits
Parental leave 22 weeks
+12
ML Performance Engineer: GPU & Systems Optimisation
ML Performance Engineer: GPU & Systems Optimisation

Trading Interview • Greater London

Hybrid
GBP 120,000 - 180,000
Senior ML Engineer — High-Performance Python & HPC
Senior ML Engineer — High-Performance Python & HPC

Longshot Systems Ltd • Greater London

Hybrid
GBP 45,000 - 65,000
Participation in company bonus scheme
10% matched pension contributions
Private healthcare insurance
+2
Research Software Engineer: Scale Distributed ML Pipelines
Research Software Engineer: Scale Distributed ML Pipelines

Deepstreamtech • Greater London

On-site
GBP 50,000 - 80,000
ML Systems Performance Engineer
ML Systems Performance Engineer

Quant Blueprint LLC • Greater London

On-site
GBP 50,000 - 70,000
Member of Technical Staff, Training Performance Engineer
Member of Technical Staff, Training Performance Engineer

Cohere • Greater London

On-site
GBP 70,000 - 90,000