Research Engineer (Agentic Models)

JetBrains

Greater London

On-site

GBP 90,000 - 140,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

JetBrains in the United Kingdom seeks a Research Engineer to own models, training loops, and evaluation pipelines powering multi-step coding agents within JetBrains IDEs.

You will work with distributed GPU clusters and design SFT/RL post-training methods, build evaluation environments, and collaborate with research, product, and infrastructure teams to ship features.

Qualifications

  • Extensive hands-on experience training LLMs in a research or production setting.
  • Deep expertise in PyTorch and specialized LLM training stacks (Megatron, NeMo, verl, or similar).
  • Strong theoretical and practical understanding of LLM fundamentals: architectures, tokenization, data pipelines, batching, mixed precision, distributed training, and debugging unstable runs.
  • The ability to own projects end to end, starting from a high-level problem or product pain point and overseeing it through the design, experimentation, implementation, and iteration phases.
  • A product-aware mindset - you care about how developers actually use agents and can translate product needs and failure modes into modeling and evaluation work.
  • At least 3 years of Python experience writing clean, maintainable code in modern ML codebases.

Responsibilities

  • Design, implement, and maintain SFT and RL post-training pipelines for multi-step coding agents.
  • Train and adapt LLMs for agent workflows, including planning, tool use, and multi-step interactions inside JetBrains IDEs.
  • Build and develop evaluation and simulation environments where coding agents can act, be measured, and compared on realistic developer tasks.
  • Design evaluation frameworks and metrics for agent behavior, analyze traces and logs, and close the loop from evaluation back into training, data, and reward design.
  • Analyze training and evaluation results to propose and implement improvements to model architectures, training recipes, and datasets.
  • Work with large-scale infrastructure, including distributed training on GPU clusters and large MapReduce-style data processing for pre-training and fine-tuning datasets.
  • Collaborate closely with research, product, and infrastructure teams to turn high-level product visions into concrete models, experiments, and shipped features.

Skills

LLM Training
PyTorch
Distributed Training
Python
Evaluation & Metrics

Tools

Kubeflow
Dagster
Airflow
Kubernetes
SLURM
ZenML

Job description

At JetBrains, code is our passion. Ever since we started, back in 2000, we've been striving to make the strongest, most effective developer tools on earth. Today, AI-powered assistance and agents are becoming a core part of how developers work in our IDEs.

We're building multi-step coding agents that can understand large codebases, plan changes, call tools, and iterate with the user. As a Research Engineer in the Agentic Models team, you'll be responsible for the models, training loops, and evaluation pipelines that power these agents.

You'll work at the intersection of SFT and RL-style post-training, and product-driven evaluation, using our distributed GPU and MapReduce clusters to ship models into JetBrains products.

As part of our team, you will:
  • Design, implement, and maintain SFT and RL post-training pipelines for multi-step coding agents.
  • Train and adapt LLMs for agent workflows, including planning, tool use, and multi-step interactions inside JetBrains IDEs.
  • Build and develop evaluation and simulation environments where coding agents can act, be measured, and compared on realistic developer tasks.
  • Design evaluation frameworks and metrics for agent behavior, analyze traces and logs, and close the loop from evaluation back into training, data, and reward design.
  • Analyze training and evaluation results to propose and implement improvements to model architectures, training recipes, and datasets.
  • Work with large-scale infrastructure, including distributed training on GPU clusters and large MapReduce-style data processing for pre-training and fine-tuning datasets.
  • Collaborate closely with research, product, and infrastructure teams to turn high-level product visions into concrete models, experiments, and shipped features.
We'll be happy to bring you on board if you have:
  • Extensive hands-on experience training LLMs (pre-training, fine-tuning, or post-training) in a research or production setting.
  • Deep expertise in modern deep learning frameworks such as PyTorch, and specialized LLM training stacks (e.g. Megatron, NeMo, verl, or similar).
  • Strong theoretical and practical understanding of LLM fundamentals: architectures, tokenization, data pipelines, batching, mixed precision, distributed training, and debugging unstable runs.
  • The ability to own projects end to end, starting from a high-level problem or product pain point and overseeing it through the design, experimentation, implementation, and iteration phases.
  • A product-aware mindset - you care about how developers actually use agents and can translate product needs and failure modes into modeling and evaluation work.
  • At least 3 years of Python experience writing clean, maintainable code in modern ML codebases.
Our ideal candidate would have experience with:
  • ML orchestrators and workflow tools such as Kubeflow, Dagster, Airflow, ZenML, and/or job schedulers like Kubernetes or SLURM.
  • Large-scale data and training pipelines, e.g. MapReduce-style clusters, multi-node GPU training, or workloads on the order of 1M+ CPU/GPU hours.
  • Designing and maintaining evaluation pipelines for LLMs or agents, including metrics, dashboards, experiment tracking, and automated regression checks.
  • AI agent development, such as tool-using agents, planners, or multi-step coding workflows, and familiarity with agentic frameworks or patterns.
  • Experiment tracking and observability using tools like Weights & Biases, MLflow, Langfuse, or similar.
  • Inference optimization and serving optimized models in production.

#LI-KP1

We are an equal opportunity employer

We know great ideas can come from anyone, anywhere. That's why we do our best to create an open and inclusive workplace - one that welcomes everyone regardless of their background, identity, religion, age, accessibility needs, or orientation.

We process the data provided in your job application in accordance with the Recruitment Privacy Policy.

Find more English Speaking Jobs in United Kingdom on Arbeitnow

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Founding ML Engineer (Spectrum)
Founding ML Engineer (Spectrum)

JetBrains • Greater London

On-site
GBP 120,000 - 160,000
JetBrains benefits
Startup autonomy
Research Engineer / Scientist, Post-training - London
Research Engineer / Scientist, Post-training - London

H Company • Greater London

On-site
GBP 60,000 - 90,000
Competitive salary
Opportunities for professional growth
Collaborative and multicultural team environment
AI Engineer
AI Engineer

G-Research • Greater London

On-site
GBP 90,000 - 150,000
Highly competitive compensation
Annual discretionary bonus
Lunch provided
+5
Member of Technical Staff (Applied AI)
Member of Technical Staff (Applied AI)

Aptura • Greater London

On-site
GBP 70,000 - 120,000
Research Engineer, Pretraining Scaling - London
Research Engineer, Pretraining Scaling - London

Anthropic • Greater London

On-site
GBP 250,000 - 435,000
Equity benefits
Visa sponsorship
Member of Technical Staff (Post Training)
Member of Technical Staff (Post Training)

Inherentlabs • Greater London

On-site
GBP 80,000 - 100,000
Research Engineer, Pretraining Scaling - London
Research Engineer, Pretraining Scaling - London

anthropic • Greater London

On-site
GBP 260,000 - 630,000
Staff AI Engineer- Developer Tools
Staff AI Engineer- Developer Tools

JetBrains • Greater London

Remote
GBP 120,000 - 180,000
Flexible work location
Remote work up to 30 days abroad
Learning and development opportunities
+1
Researcher, Training - London
Researcher, Training - London

United States Digital Space LLC • Greater London

Hybrid
GBP 70,000 - 90,000
Relocation support
Hybrid work schedule
Senior AI Engineer
Senior AI Engineer

Profitero, Inc. • Wokingham

On-site
GBP 70,000 - 90,000
Competitive base salary
Employee healthcare
25 days off + bank holidays
+1