Senior ML Engineer _TT

PulseRise Technologies

Greater London

On-site

GBP 120,000 - 180,000

Full time

9 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

PulseRise Technologies is hiring a Senior ML Engineer in London to own the end-to-end design, training, and deployment of a novel foundation model. You will create custom CUDA kernels, scale distributed training on cloud infra, and develop tooling for a fast-moving team.

This hands-on role demands leadership and rapid delivery in an early-stage environment. You will work directly with the founders, shaping the model and infrastructure to meet production SLAs while maintaining high standards for

Qualifications

  • Shipped a large-scale foundation model in a high-growth AI/ML setting.
  • Designed and implemented custom CUDA kernels for model optimization.
  • Direct experience scaling distributed training and/or inference on cloud infra.
  • Deep knowledge of PyTorch and recent architectures (e.g., MoE, state-space models).
  • Hands-on ML systems ownership with production SLAs and tooling experience.
  • Proven ability to deliver quickly in ambiguous, fast-paced enviroments.
  • Fluent English communication and collaboration with founders.

Responsibilities

  • Architect and implement a large-scale foundation model from research to production.
  • Design and write custom CUDA kernels to optimize performance where libraries fall short.
  • Build and scale distributed training and inference pipelines on cloud infrastructure.
  • Profile, debug, and optimize DL models for latency and throughput against SLAs.
  • Build internal tooling and infrastructure to accelerate iteration speed.
  • Partner with founders to set technical direction in a fast-paced, ambiguous setting.
  • Take hands-on technical leadership as the team scales.

Skills

CUDA C/C++
Python
PyTorch
Distributed training
Model optimization
Production ML systems
English fluency

Tools

AWS
Azure
GCP
CUDA profiling tools

Job description

We're hiring a Senior ML Engineer to own the design, training, and deployment of a novel foundation model — from research through production, including the custom CUDA kernels that make it fast. This is a hands‑on, high‑ownership role for someone who has already shipped a large-scale foundation model (01) at a high‑growth AI/ML startup or top‑tier research lab, not just published about one. You'll architect and scale distributed training and inference pipelines on cloud infrastructure, profile and optimize deep learning models at the systems level, and build the internal tooling that lets a small, fast-moving team punch above its size. The environment is early‑stage, high‑transparency, and high‑urgency — decisions move quickly, ambiguity is the norm, and the team expects people to challenge and be challenged. You'll work closely with the founders against real production SLOs and SLAs rather than research benchmarks. If you want to be the hands‑on technical owner of a first‑of‑its‑kind product rather than one contributor among many, this role is built for you.

Details
  • Schedule: Full-time
  • Location: UK, London
  • Start: ASAP
  • Duration: Long-term
  • English: Fluent
  • Type of collaboration: B2B
About The Project

The client is a VC‑backed AI/ML startup building a novel foundation model that enables fully automated, unsupervised software delivery for embedded control systems. It's an early‑stage company at a critical growth point, scaling its technical team to deliver a high‑impact, first‑of‑its‑kind product. The culture is direct, high‑transparency, and no‑jargon — the team values honesty, urgency, and strong work ethic over process and hierarchy. Technical leadership is hands‑on and expects the same from every hire: this is not a role for someone who wants to hand off hard problems to others. The company operates with real ambiguity and rapid change, and rewards people who take ownership and move fast. Candidates should be excited by the prospect of building something genuinely novel from the ground up, not maintaining an existing system.

You have
  • Shipped a large-scale foundation model (01) at a high-growth AI/ML startup or top-tier research lab — hands‑on delivery, not purely academic or research‑only experience
  • Designed and implemented custom CUDA kernels for model optimization, with strong proficiency in both CUDA C/C++ and Python
  • Direct experience scaling distributed training and/or inference pipelines on cloud infrastructure (AWS, Azure, or GCP)
  • Deep knowledge of at least one major deep learning framework, ideally PyTorch, and hands‑on experience with recent architectures (e.g., MoE, state‑space models)
  • Hands‑on ownership of ML systems with strict SLOs or production SLAs — you've operated systems in production, not just built models
  • A track record of building internal tooling or infrastructure that measurably accelerated a team's productivity
  • Demonstrated ability to deliver quickly in ambiguous, fast‑paced, early‑stage environments
  • Fluent English
What To Do
  • Architect and implement a large‑scale foundation model, from research through production deployment
  • Design and write custom CUDA kernels to optimize model performance where off‑the‑shelf libraries fall short
  • Build and scale distributed training and inference pipelines on cloud infrastructure
  • Profile, debug, and optimize deep learning models for latency, throughput, and reliability against production SLAs
  • Build internal tooling and infrastructure to accelerate the team's iteration speed
  • Work directly with the founders to make fast, high‑ownership technical decisions in an ambiguous, high‑urgency environment
  • Take hands‑on technical leadership as the team scales, helping set technical direction for the model and its infrastructure
Interview Process
  • 1‑hour online cultural interview with the CEO (focus: values, urgency, transparency, team fit)
  • 2‑hour technical interview with the CPO — deep technical deep‑dive, hands‑on problem‑solving, system design. No live or take‑home coding tasks.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior ML Engineer
Senior ML Engineer

TechTree • Greater London

On-site
GBP 120,000 - 180,000
Founding Engineer | ML & Data Science | London Based Start-Up
Founding Engineer | ML & Data Science | London Based Start-Up

Oho Group • Greater London

On-site
GBP 90,000 - 130,000
Founding Machine Learning Research Engineer
Founding Machine Learning Research Engineer

Generative • Greater London

On-site
GBP 120,000 - 160,000
Senior ML Infrastructure Engineer (Research Initiatives) - Systems Integrator
Senior ML Infrastructure Engineer (Research Initiatives) - Systems Integrator

Hamilton Barnes Associates Limited • United Kingdom

On-site
GBP 90,000 - 130,000
Significant stock option packages
Remote-first working setup
Fully paid travel and accommodation
+1
Foundation Model Engineer
Foundation Model Engineer

Pure Resourcing Solutions Limited • Cambridge

On-site
GBP 90,000 - 140,000
Founding AI Engineer
Founding AI Engineer

X4 Engineering • Greater London

On-site
GBP 70,000 - 120,000
Equity
Annual bonus
Lunch & dinner
+2
Lead Foundation ML Engineer – Architecture & Deployment
Lead Foundation ML Engineer – Architecture & Deployment

TechTree • Greater London

On-site
GBP 120,000 - 180,000
Staff AI Engineer
Staff AI Engineer

United States Digital Space LLC • Greater London

Remote
GBP 114,000 - 159,000
Member of Technical Staff
Member of Technical Staff

Salient Group • Greater London

Hybrid
GBP 120,000 - 180,000
Equity grant
Full-stack Engineer London
Full-stack Engineer London

Model ML • Greater London

On-site
GBP 70,000 - 115,000
Equity
Competitive salary