Member of Technical Staff (Post Training)

Inherentlabs

Greater London

On-site

GBP 80,000 - 100,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Inherentlabs in London is seeking Members of Technical Staff to lead post-training state-of-the-art foundation models for scientific research. You will design algorithms, run experiments, and optimize workflows in a team-oriented environment.

Ideal candidates have strong deep learning and software engineering backgrounds, with a passion for redefining AI capabilities. If you resonate with our mission to explore unknown knowledge, we invite you to apply and contribute to shaping the AI frontier.

Qualifications

  • 3+ years of deep learning research experience.
  • 5+ years of software engineering experience.
  • Experience post-training large language, vision, video or multi-modal models.

Responsibilities

  • Design, implement, and tune SFT and RL algorithms.
  • Build autocurricula and reward signals for research.
  • Run large-scale experiments on state-of-the-art hardware.

Skills

Deep learning research
Software engineering
Python
Deep learning frameworks (e.g., PyTorch)

Education

PhD in mathematics, computer science or hard science discipline

Job description

Member of Technical Staff, Post-Training — Inherent (London)

At Inherent, we are on a mission to build AI that recursively self-improves to discover new knowledge. Scientific advances are the backbone of our economic, technological and societal prosperity, but ideas are getting harder to find and breakthroughs are becoming more expensive. We are building a new frontier lab dedicated to developing AI that explores “unknown unknowns” to uncover paradigm-shifting research contributions. Science is a social endeavour, and so our mission is inextricably a human-machine teaming problem. We’re starting by reinventing the AI research factory so that our own agents accelerate their own creation.

Inherent is a well-funded, fast-growing neo-lab backed by Tier 1 VCs who believe in our ethical stance. We are a team of operators with backgrounds at frontier labs who have done foundational work in recursive self-improvement, AI Scientists, world modelling, meta-RL and human-machine cooperation. Working in-person every day at our high-intensity London headquarters, we believe that Europe will lead the way in the coming paradigm of AI-enabled science, unlocking human potential across the globe.

About the role

We’re looking for Members of Technical Staff to lead work on post-training state-of-the‑art foundation models for open-ended agentic capabilities in scientific research. You’ll be involved at every level of the post-training pipeline: sourcing and creating data, building autocurricula, devising and implementing SFT and RL algorithms, constructing tools and harnesses for foundation model self‑improvement, analysing research results, and using information gained to devise future hypotheses. You will work closely with an experienced technical team of humans, and increasingly alongside the AI scientist collaborators we dogfood.

What you'd do
  • Design, implement, and tune SFT and RL algorithms to post-train models that autonomously perform state-of-the‑art research.
  • Build the autocurricula, judges, harnesses and eval pipelines that turn open-ended research tasks into reliable reward signal.
  • Run large-scale experiments on state-of-the‑art hardware and analyse experiments to determine the next hypotheses to test, in collaboration with our AI agents.
  • Close recursive loops so that AI agents drive their own post-training research.
  • Work closely with colleagues in the Infrastructure and AI for Science teams to optimise hardware and deliver remarkable performance in real scientific domains.
What we're looking for
  • 3+ years of deep learning research experience.
  • Experience post-training large language, vision, video or multi-modal models.
  • Demonstrated track record of success in deep learning research, whether papers, model releases, open-source contributions, or other artifacts.
  • 5+ years of software engineering experience, including deep familiarity with Python and at least one deep learning framework (e.g., PyTorch, JAX).
  • Experience using the latest coding agents, and opinions about optimal workflow.
  • Enthusiasm for experimental organizational design.
  • AI-pilled: adopting agents, keen to build a company where agents are front and centre.
Strong candidates may also have
  • PhD in mathematics, computer science or hard science discipline.
  • Hands‑on experience training LLMs with RL at scale (GRPO/PPO, DPO, distillation, and variants).
  • Familiarity with distributed and long-context training infrastructure.
  • A background in autocurricula, open‑endedness, meta‑learning, or recursive self‑improvement.
  • Experience post‑training frontier models at an industry lab (scale, infra, and iteration speed).
Why this is interesting
  • You’ll shape the core research of a frontier AI lab from the beginning.
  • You’ll work on genuine recursive self‑improvement — training AI scientists that improve the very pipeline that trains them — not incremental benchmark‑chasing.
  • You’ll dogfood your own work: the agents you post‑train accelerate the research that creates them.
  • Small team, high trust, no bureaucracy, and a genuinely technical culture.
Culture

We only select people with low ego, spiky skill profiles, commitment to societal benefit, unusual viewpoints, and a passion for "living in the experiment". We'll win because we're willing to try things that no incumbent would even think to do, let alone action.

We have really good lunch and dinner. Seriously. You've got to try it. We're based in King's Cross, London and believe in the pace and energy of working in person. We’re committed to having the most tasteful, and the weirdest, office of any AI lab: the environment shapes the agents within it.

If you believe in our mission and culture, and are qualified and motivated, we encourage you to apply, even if you don’t meet every one of the criteria above. We know that many of the most creative and talented people have had unusual career paths and backgrounds. Building a team with a diversity of thought is mission‑critical, for plurality spurs curiosity, invention and collective experimentation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff (AI for Science)
Member of Technical Staff (AI for Science)

Inherentlabs • Greater London

On-site
GBP 65,000 - 85,000
Member of Technical Staff (Infrastructure Engineer, Training and Inference Systems)
Member of Technical Staff (Infrastructure Engineer, Training and Inference Systems)

Inherentlabs • Greater London

On-site
GBP 70,000 - 90,000
Member of Technical Staff (Infrastructure Engineer, Compute Infrastructure)
Member of Technical Staff (Infrastructure Engineer, Compute Infrastructure)

Inherentlabs • Greater London

On-site
GBP 70,000 - 90,000
Good lunch and dinner
Collaborative work culture
No bureaucracy
Member of Technical Staff (Factory Redesign)
Member of Technical Staff (Factory Redesign)

Inherentlabs • Greater London

On-site
GBP 60,000 - 80,000
Great office culture
Good lunch and dinner
Research Engineer / Scientist, Post-training - London
Research Engineer / Scientist, Post-training - London

H Company • Greater London

Hybrid
GBP 60,000 - 90,000
Competitive salary
Opportunities for professional growth
Collaborative and multicultural team environment
Research Engineer / Scientist, Post-training - London
Research Engineer / Scientist, Post-training - London

H Company • Greater London

Hybrid
GBP 120,000 - 180,000
Senior Research Engineer
Senior Research Engineer

Basecamp Research • Greater London

On-site
GBP 90,000 - 125,000
Impactful Mission
Collaborative Culture
High Growth
+1
Research Engineer - Model Ablation
Research Engineer - Model Ablation

Ellison Institute of Technology • Oxford

On-site
GBP 70,000 - 120,000
Enhanced holiday pay
Pension
Life Assurance
+6
Research Engineer, RL Scaling Science New London, UK
Research Engineer, RL Scaling Science New London, UK

Alcides Fonseca • Greater London

Hybrid
GBP 70,000 - 90,000
Competitive compensation
Generous vacation and parental leave
Flexible working hours
Research Engineer - Model Ablation
Research Engineer - Model Ablation

Ellison Institute of Technology Oxford • Oxford

On-site
GBP 90,000 - 120,000
Enhanced holiday pay
Pension
Life Assurance
+6