Production ML Researcher for Finance - Low-Latency AI

Thurn Partners Ltd

New York (NY)

On-site

USD 180,000 - 300,000

Full time

16 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Thurn Partners Ltd in New York seeks a post-training ML researcher to adapt large pretrained models for financial markets, handling noisy data and distributions that shift daily while meeting tight latency budgets.

You will work with production teams to deploy models, evaluate against real market outcomes, and optimize GPU performance with C++/Python, CUDA and bespoke kernels for low-latency inference.

Qualifications

  • Degree in quantitative field with GPA ≥ 3.5/4.0
  • Industry ML experience with hands-on post-training work
  • Production-quality Python and C++
  • GPU knowledge of memory hierarchy and how to diagnose a slow kernel
  • Eligibility to work in the US

Responsibilities

  • Fine-tune, distil and align large models for domain-specific tasks
  • Build evaluation that measures what matters, not what is easy to measure
  • Quantize, prune and optimize inference where latency carries a direct cost
  • Write and profile custom GPU kernels when standard ops become the bottleneck
  • Take research into production with engineering and trading teams

Skills

ML experience
Leadership experience

Education

Bachelor's degree in quantitative field (GPA ≥ 3.5/4.0)
Master's or PhD in ML/CS/Math/Stats/Physics

Tools

Python
C++
CUDA
Triton
Nsight
NCCL

Job description

Thurn Partners Ltd in New York seeks a post-training ML researcher to adapt large pretrained models for financial markets, handling noisy data and distributions that shift daily while meeting tight latency budgets.

You will work with production teams to deploy models, evaluate against real market outcomes, and optimize GPU performance with C++/Python, CUDA and bespoke kernels for low-latency inference.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Production ML Researcher: Latency-Sensitive Finance Models
Production ML Researcher: Latency-Sensitive Finance Models

Thurn Partners Ltd • Miami (FL)

On-site
USD 150,000 - 230,000
Senior ML Researcher — Real-Time Trading Models that Move P&L
Senior ML Researcher — Real-Time Trading Models that Move P&L

Thurn Partners Ltd • New York (NY)

On-site
USD 180,000 - 300,000
Senior Machine Learning Researcher, Post-Training & Inference Optimization
Senior Machine Learning Researcher, Post-Training & Inference Optimization

Thurn Partners Ltd • New York (NY)

On-site
USD 180,000 - 300,000
Senior Machine Learning Researcher, Post-Training & Inference Optimization
Senior Machine Learning Researcher, Post-Training & Inference Optimization

Thurn Partners Ltd • Miami (FL)

On-site
USD 150,000 - 230,000
AI Research Engineer — Production-Scale ML & HPC
AI Research Engineer — Production-Scale ML & HPC

Jump Trading • New York (NY)

On-site
USD 270,000 - 360,000
High-Performance ML Inference Engineer for Markets
High-Performance ML Inference Engineer for Markets

Fintal Partners • New York (NY)

On-site
USD 140,000 - 220,000
Low-Latency ML Inference Engineer
Low-Latency ML Inference Engineer

Career Techniques • New York (NY)

Hybrid
USD 200,000 - 300,000
Senior ML Researcher - Quant Trading, P&L Impact
Senior ML Researcher - Quant Trading, P&L Impact

Thurn Partners Ltd • United States

On-site
USD 180,000 - 260,000
Deep Learning ML Research Intern – Fast-Paced FinTech
Deep Learning ML Research Intern – Fast-Paced FinTech

Jump Trading • New York (NY)

On-site
USD 260,000 - 360,000
AI Research Engineer Intern — Scalable Finance ML
AI Research Engineer Intern — Scalable Finance ML

Trading Interview • New York (NY)

On-site
USD 255,000 - 345,000