AI Research Engineer

Arbitr Group Ltd

Philadelphia (Philadelphia County)

Hybrid

USD 140,000 - 210,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Hybrid work model
Competitive benefits

Job summary

Arbitr Group Ltd is seeking an AI Research Engineer to build models that ship and the surrounding systems, from data pipelines to production deployment.

You will own model construction, retrieval, memory, evaluation, and collaboration with product and partnerships while working in a hybrid Philadelphia setting.

Qualifications

  • Experience building and training ML models end-to-end, from data processing to production serving.
  • Strong Python skills and ability to debug at system level.
  • Familiarity with model evaluation, benchmarks, and experimental governance.

Responsibilities

  • Build custom models end-to-end for customers and internal platforms.
  • Develop systems around models: retrieval, memory, reasoning, orchestration.
  • Own evaluation tooling and benchmarks to measure true model quality.
  • Push research into production and track new models/techniques.

Skills

Python
Model building
Reinforcement learning
Dataset design

Tools

HuggingFace

Job description

AI Research Engineer

Do research that ships, and build the systems around it.

arbitr isn't asking you to fine-tune someone else's roadmap. We're asking you to help define it.

We build models that run inside our platforms. We build the retrieval, memory, and evaluation systems around them, and we ship the whole thing into production, where it has to actually work.

arbitr is a 25-year global company (formerly Straker Ltd) — ASX-listed, offices across the USA, Japan, Europe and New Zealand - that has always been at the forefront of technology, and that includes AI, from the inside out. We operate in smaller teams, as a startup, which means you get something rare: the scale and stability of an established business, with the mandate and the appetite for risk of a team that is genuinely starting again.

If you want to do research that ships, not research that lives in a notebook, this is that role.

The team

You’ll join our AI research and engineering team, reporting to and collaborating with the Head of AI Research and Model Engineering. We build the models that power arbitr’s platforms, we build the systems that give those models the right context and structure to be useful, and we act as the AI experts across the company and for our clients, providing technical guidance to product, engineering, and partnerships.

What you’ll own
  • Build custom models. For customers, and across arbitr's platforms. You own the pipeline end-to-end - first experiments through production and beyond.
  • Build the systems around the models. Retrieval. Memory. Reasoning. Orchestration. Whatever the model needs on either side of it to actually deliver value in production.
  • Own our evals. Build the benchmark tooling that tells us whether a model is genuinely good, not just good on paper.
  • Push the research. Track new models and training techniques as the field evolves. Bring the ones that matter into our stack.
  • Design the research workstreams. Take an idea from a problem statement, through experiments, findings, and the pipeline changes on the other side.
  • Deploy and run models at scale. Serve, monitor, and orchestrate models on our GPU infrastructure. Production is your problem to own.
  • Work at the partner table. Support strategic engagements with global partners, where your technical judgement carries weight.
  • Go wide across arbitr. Bring engineering support to arbitr's wider platform work: architecture reviews, AI strategy input and hands‑on development where it counts.
You’ll thrive here if

You’ve trained a model before. Not fine-tuning someone else's on HuggingFace, but actually training from the ground up. You understand what's happening under the hood and you can debug when it goes wrong.

You think in experiments, not vibes. You build real evaluation gates before you trust a result.

You are relentless about getting from idea to production. You filter hard for what is signal versus noise. You want to join a team in the middle of reinventing itself, and leave your work in the shape of what comes next.

What we’re looking for
  • Detailed understanding of model building - concepts like LoRA, PT, reinforcement learning, and inference are familiar territory.
  • Comfort designing and building datasets, and a working background in data science.
  • Ability to work alone on a project, or with a team of like minded individuals.
  • Strong Python skills.
  • A self-starter mindset. You are part of a team, but personally accountable for outcomes.
  • Familiarity with the ideas we write about at labs.straker.ai, or the appetite to get familiar quickly.
Practical details
  • Location:Philadelphia/ hybrid
  • Work rights: Must have working rights for USA
  • Compensation: competitive base + benefits
  • Diversity: we hire for judgement, curiosity and evidence, not for pedigree. We actively welcome applicants from backgrounds under-represented in AI research.
  • Accessibility: if you need accommodations at any stage of the interview process, tell us and we’ll make them.
Why arbitr, why now

25 years of proven results. A global footprint across the USA, Japan, Europe and New Zealand. And a company that has just renamed itself around the work it now does - with a new mandate to push to be AI-native from the ground up.

We're building the infrastructure, research systems and shared layers to move with a field that changes month to month. If you want to help shape what that foundation looks like, we'd like to hear from you.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Product Engineer
Product Engineer

Arbitr Group Ltd • Philadelphia

On-site
USD 110,000 - 170,000
Account Executive
Account Executive

Arbitr Group Ltd • Philadelphia

On-site
USD 90,000 - 160,000
Product Solutions Engineer
Product Solutions Engineer

Arena • San Francisco (CA)

On-site
USD 160,000 - 210,000
Equity
Health benefits
AI at work
+1
AI Solutions Engineer — Customer Relations & Platform Build
AI Solutions Engineer — Customer Relations & Platform Build

Arena • San Francisco (CA)

On-site
USD 160,000 - 210,000
Research Engineer (Production Model Post Training)
Research Engineer (Production Model Post Training)

Menlo Ventures • San Francisco (CA)

On-site
USD 315,000 - 510,000
Flexible working hours
Generous vacation and parental leave
Office space for collaboration
Applied AI, Research Engineer
Applied AI, Research Engineer

Anthropic • San Francisco (CA)

On-site
USD 300,000 - 400,000
Competitive compensation
Equity donation matching
Generous vacation & parental leave
+2
Staff+ Software Engineer, RL Data Platform San Francisco, CA | New York City, NY
Staff+ Software Engineer, RL Data Platform San Francisco, CA | New York City, NY

Anthropic Limited • San Francisco (CA), Northern (KY)

On-site
USD 320,000 - 405,000
Equity donation matching
Generous vacation & parental leave
Flexible working hours
Research Engineer, RL Engineering
Research Engineer, RL Engineering

Anthropic • New York (NY)

On-site
USD 500,000 - 850,000
Equity donation matching
Generous vacation
Parental leave
+2
Research Engineer, RL Engineering
Research Engineer, RL Engineering

Anthropic • San Francisco (CA)

On-site
USD 500,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation and parental leave
+1
Applied AI, Research Engineer
Applied AI, Research Engineer

Anthropic • New York (NY)

On-site
USD 300,000 - 400,000
Competitive compensation
Equity donation matching
Generous vacation
+3