Staff AI Engineer

United States Digital Space LLC

Greater London

Remote

GBP 114,000 - 159,000

Full time

7 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

United States Digital Space LLC is seeking a Staff Engineer for AI/ML to own the end-to-end self-serve adaptation toolchain. You will enable customers to adapt models on private data without exposing it externally, covering data processing, labeling, QA, active learning, training, evaluation, and model promotion, with on-site deployment options.

You will build reproducible pipelines, versioned datasets, and evaluation methods while leading a small team focusing on curation, auto-labeling,

Qualifications

  • 5+ years building production ML, AI, or computer-vision systems.
  • Strong Python and PyTorch.
  • Experience owning ML data pipelines, training pipelines, or evaluation infrastructure.
  • Deep CV data experience: detection, segmentation, annotation taxonomies, dataset curation, and data quality.
  • Hands-on experience with model-in-the-loop or foundation-model-assisted labeling.
  • Familiarity with tools such as SAM-2, Grounding DINO, FiftyOne, and annotation platforms.
  • Strong understanding of evaluation: held-out sets, leakage prevention, baselines, slice metrics, and promotion gates.
  • Experience with QA sampling, label-noise analysis, missing annotations, IAA, or alternatives when only one operator is available.
  • Experience with active learning or other methods for prioritizing labeling effort.
  • Ability to reason about dataset economics: quality vs. quantity, long-tail coverage, and cost-per-useful-example.
  • Experience with dataset versioning, lineage, experiment tracking, model registries, or data cards.
  • Strong ownership, communication, and systems thinking.
  • Experience leading engineers as a tech lead, staff engineer, or small-team manager.

Responsibilities

  • Own the ML adaptation pipeline from raw customer data to trained, evaluated, deployable models.
  • Build foundation-model-assisted labeling workflows using tools such as SAM-2, Grounding DINO, open-vocabulary models, LLM steering, and human review.
  • Design self-serve operator workflows using yes/no/maybe feedback and natural-language corrections.
  • Create versioned datasets with lineage, data cards, label provenance, class distributions, and known gaps.
  • Build QA methods that catch systematic pseudo-label errors, missing annotations, and long-tail data gaps.
  • Develop active-learning loops that prioritize the highest-value frames for limited operator review.
  • Build reproducible train/eval pipelines with experiment tracking, model packaging, and promotion gates.
  • Design evaluation around held-out anchor sets, leakage prevention, baselines, slice metrics, and automated approve/reject decisions.
  • Package the toolchain for on-prem, air-gapped, regulated, or customer-held environments.
  • Close the field-failure loop by feeding live failures back into the next tune cycle.
  • Lead and grow a small specialist team around curation, auto-labeling, deployment, and onboarding.

Skills

Python
PyTorch
Computer vision
ML data pipelines
Model-in-the-loop
Active learning
Grounding DINO
SAM-2
FiftyOne
Annotation tools
Leadership
Systems thinking

Tools

SAM-2
Grounding DINO
FiftyOne
Annotation platforms

Job description

the company exists to make aggression unaffordable.

Founded in Ukraine and headquartered in London, the company is a defence-technology company with operations across Europe, the United States and Asia.

the company builds uncrewed vessels, ground robotics and aircraft, autonomy software, and the command-and-control layer that operates them together. Each platform is deployable on its own, and stronger as part of the system. Its product lines include the MAGURA family of uncrewed surface vessels, the NEMESIS family of strike platforms, the LIUT family of ground robotic platforms, and counter-UAS systems.

About the role

The Staff Engineer, AI / ML — Self-Serve Toolchain will build the end-to-end system that lets customers adapt the company models on their own private data without exposing that data to us.

This role spans data processing, foundation-model-assisted labeling, human-in-the-loop QA, active learning, training, evaluation, and model promotion. You will turn an in-flight principal-led capability into a repeatable toolchain that non-expert customers can run safely on-site.

What you’ll do

Own the ML adaptation pipeline from raw customer data to trained, evaluated, deployable models.

  • Build foundation-model-assisted labeling workflows using tools such as SAM-2, Grounding DINO, open-vocabulary models, LLM steering, and human review.
  • Design self-serve operator workflows using yes/no/maybe feedback and natural-language corrections.
  • Create versioned datasets with lineage, data cards, label provenance, class distributions, and known gaps.
  • Build QA methods that catch systematic pseudo-label errors, missing annotations, and long-tail data gaps.
  • Develop active-learning loops that prioritize the highest-value frames for limited operator review.
  • Build reproducible train/eval pipelines with experiment tracking, model packaging, and promotion gates.
  • Design evaluation around held-out anchor sets, leakage prevention, baselines, slice metrics, and automated approve/reject decisions.
  • Package the toolchain for on-prem, air-gapped, regulated, or customer-held environments.
  • Close the field-failure loop by feeding live failures back into the next tune cycle.
  • Lead and grow a small specialist team around curation, auto-labeling, deployment, and onboarding.
What success looks like

Customers can run a full adaptation cycle without engineer intervention.

  • Customer data stays inside the customer boundary.
  • The system produces trusted datasets with clear provenance and quality signals.
  • Pseudo-label quality is measured and systematic errors are caught early.
  • The promotion gate can approve or reject models based on evidence, not intuition.
  • Evaluation is protected by anchor sets, leakage controls, slice metrics, and baseline comparisons.
  • Field failures become reproducible inputs to the next training cycle.
  • Synthetic data is used only when it proves value against real held-out data.
  • The toolchain becomes a repeatable capability supported by a small, ramped team.
Required Qualifications

5+ years building production ML, AI, or computer-vision systems.

  • Strong Python and PyTorch.
  • Experience owning ML data pipelines, training pipelines, or evaluation infrastructure.
  • Deep CV data experience: detection, segmentation, annotation taxonomies, dataset curation, and data quality.
  • Hands-on experience with model-in-the-loop or foundation-model-assisted labeling.
  • Familiarity with tools such as SAM-2, Grounding DINO, FiftyOne, and annotation platforms.
  • Strong understanding of evaluation: held-out sets, leakage prevention, baselines, slice metrics, and promotion gates.
  • Experience with QA sampling, label-noise analysis, missing annotations, IAA, or alternatives when only one operator is available.
  • Experience with active learning or other methods for prioritizing labeling effort.
  • Ability to reason about dataset economics: quality vs. quantity, long-tail coverage, and cost-per-useful-example.
  • Experience with dataset versioning, lineage, experiment tracking, model registries, or data cards.
  • Strong ownership, communication, and systems thinking.
  • Experience leading engineers as a tech lead, staff engineer, or small-team manager.
Nice to have

Synthetic data, sim2real, or domain randomization experience.

  • EO / IR / LWIR, remote sensing, maritime imagery, or small-object detection experience.
  • Privacy-preserving ML, federated learning, on-prem, or air-gapped deployment experience.
  • Experience building self-serve ML platforms or tools for non-expert users.
  • Experience with lakeFS, DVC, MLflow, Weights & Biases, Kubernetes, Kubeflow, Flyte, Dagster, Airflow, or Argo.
  • LLM application, context engineering, structured output, or LLM evaluation experience.
  • Exposure to radar, AIS, EO/IR fusion, tracking, sensor fusion, robotics, autonomy, UxV, defence tech, C2/C4ISR, or tactical systems.

The nature of combat has changed.

Tomorrow’s battlefield success depends on autonomy, speed, and adaptability.

And the company is ready.

Are You?

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff AI Engineer
Staff AI Engineer

Uforce • Greater London

Remote
GBP 120,000 - 190,000
Applied AI/ML Engineer for Production Deployments
Applied AI/ML Engineer for Production Deployments

brainco • Greater London

On-site
GBP 90,000 - 130,000
Competitive salary
Paid maternity and paternity leave
Daily lunches
+2
Principal ML Platform Engineer
Principal ML Platform Engineer

Synthesia • Greater London

On-site
GBP 70,000 - 100,000
Senior Machine Learning Engineer (Safety)
Senior Machine Learning Engineer (Safety)

Faculty • Greater London

On-site
GBP 90,000 - 150,000
ML Engineer (Forward Deployed)
ML Engineer (Forward Deployed)

Applied Computing • Greater London

On-site
GBP 90,000 - 130,000
Senior AI Solutions Engineer
Senior AI Solutions Engineer

Ocho People • Belfast City District

On-site
GBP 90,000 - 120,000
Staff+ Software Engineer (RL Data Platform)
Staff+ Software Engineer (RL Data Platform)

Anthropic • York and North Yorkshire

On-site
GBP 90,000 - 130,000
Health insurance
Equity options
Relocation support
+5
Senior Forward Deployed Engineer
Senior Forward Deployed Engineer

Faculty Science Limited • Greater London

On-site
GBP 90,000 - 130,000
Unlimited Annual Leave Policy
Private healthcare and dental
Enhanced parental leave
+3
Research Engineer, Pretraining Scaling - London
Research Engineer, Pretraining Scaling - London

Mat Vin • Greater London

On-site
GBP 260,000 - 630,000
Senior Software Engineer (Safety)
Senior Software Engineer (Safety)

Faculty • Greater London

On-site
GBP 90,000 - 120,000