Member of Technical Staff, Backend - NomadicML

Praxis, Inc.

San Francisco (CA)

On-site

USD 170,000 - 250,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

NomadicML in San Francisco is seeking a Backend / Infrastructure Engineer to join our platform that powers video intelligence. You’ll build scalable cloud ingestion, distributed GPU inference pipelines, and robust APIs/SDKs used by enterprises worldwide.

You’ll collaborate with ML researchers to productionize models, automate deployment, and improve reliability, speed, and developer experience across storage, scheduling, and orchestration.

Qualifications

  • Backend systems and cloud experience for large-scale inference.
  • Experience designing REST/gRPC APIs and developer-facing SDKs.
  • Knowledge of asynchronous job orchestration frameworks.
  • Hands-on with GPU inference scaling and distributed compute.
  • Ability to productionize research-grade systems.

Responsibilities

  • Build and scale the backbone powering NomadicML's video intelligence platform.
  • Collaborate with ML researchers to productionize models and automate deployment.
  • Expose capabilities via clean APIs/SDKs and improve observability.

Skills

Python
Go
TypeScript

Tools

AWS
GCP
Azure
Kubernetes
Docker
Ray
Dagster
Temporal

Job description

About NomadicML

Americans drive over 5 trillion miles a year, more than 500 billion of them recorded. Buried in that footage is the next frontier of machine intelligence. At NomadicML, we’re building the platform that unlocks it.

Our Vision-Language Models (VLMs) act as the new “hydraulic mining” for video, transforming raw footage into structured intelligence that powers real-world autonomy and robotics. We partner with industry leaders across self-driving, robotics, and industrial automation to mine insights from petabytes of data that were once unusable.

NomadicML was founded by Mustafa Bal and Varun Krishnan, who met at Harvard University while studying Computer Science.

  • Mustafa is a core contributor to ONNX Runtime and DeepSpeed with deep expertise in distributed systems and large-scale model training infrastructure

  • Varun is an INFORMS Wagner Prize Finalist for his research in large-scale driver navigation AI models and one of the top chess players in the US.

Our team has built mission-critical AI systems at Snowflake, Lyft, Microsoft, Amazon, and IBM Research, holds top-tier publications in VLMS and AI at conferences like CVPR, and moves with the speed and clarity of a startup obsessed with impact.

About the Role

We’re looking for a Backend / Infrastructure Engineer who thrives at the intersection of cloud systems, SDK design, and large-scale inference infrastructure.

You’ll build and scale the backbone that powers NomadicML’s video intelligence platform — from secure cloud ingestion to distributed GPU inference pipelines that run our largest foundation models. You’ll collaborate with ML researchers to productionize their models, automate deployment and scaling, and expose those capabilities through clean APIs and SDKs used by enterprises worldwide.

This role blends systems engineering, distributed compute orchestration, and developer experience. You’ll be working across cloud storage, inference scheduling, GPU clusters, and the NomadicML SDK.

You Might Be a Fit If You Have
  • Deep proficiency in Python, Go, or TypeScript for backend systems.

  • Experience with AWS, GCP, or Azure (IAM, S3/Blob Storage, Batch/Compute APIs, etc.).

  • Strong understanding of GPU inference scaling, Kubernetes, container orchestration, and event-driven pipelines.

  • Prior experience designing REST/gRPC APIs, SDKs, or developer-facing infrastructure.

  • Familiarity with asynchronous job orchestration (Ray, Airflow, Dagster, Temporal).

  • A practical mindset: you take research-grade systems and make them reliable, fast, and usable.

Nice to Have
  • Experience contributing to inference orchestration frameworks or ML infra tools (e.g., DeepSpeed, Triton, Ray Serve).

  • Understanding of video encoding, chunking, and streaming formats for efficient multi-modal ingestion.

  • Basic front-end experience (React / Next.js) for integrating backend pipelines into product workflows.

  • Background in ML infrastructure, observability, or data management systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff, Backend
Member of Technical Staff, Backend

Nomadic AI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Member of Technical Staff, Machine Learning
Member of Technical Staff, Machine Learning

Nomadic AI • San Francisco (CA)

On-site
USD 170,000 - 250,000
Member of Technical Staff, Machine Learning - NomadicML
Member of Technical Staff, Machine Learning - NomadicML

Praxis, Inc. • San Francisco (CA)

On-site
USD 150,000 - 240,000
Chief of Staff - NomadicML
Chief of Staff - NomadicML

Praxis, Inc. • San Francisco (CA)

On-site
USD 120,000 - 190,000
Member of Technical Staff, Frontend
Member of Technical Staff, Frontend

Praxis, Inc. • San Francisco (CA)

On-site
USD 140,000 - 190,000
Member of Technical Staff, Frontend
Member of Technical Staff, Frontend

Pear VC • Austin (TX)

On-site
USD 80,000 - 120,000
Chief of Staff
Chief of Staff

Nomadic AI • San Francisco (CA)

On-site
USD 150,000 - 210,000
Backend/Infra Engineer for Scalable Video AI Platform
Backend/Infra Engineer for Scalable Video AI Platform

Nomadic AI • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
UI/UX Designer
UI/UX Designer

Praxis, Inc. • San Francisco (CA)

On-site
USD 90,000 - 150,000
UI/UX Designer
UI/UX Designer

Nomadic AI • San Francisco (CA)

Hybrid
USD 100,000 - 160,000