Backend AI Platform Engineer – Scalable ML Infra

Engg

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Nexxa is seeking a Backend AI Engineer to design, build, and own core AI infrastructure powering production GenAI, ML, and CV solutions. You will create model-serving pipelines, inference services, data pipelines, and APIs to connect models with enterprise data sources.

You’ll architect scalable, real-time and batch systems, implement RAG pipelines, and drive reliability, observability, and CI/CD for ML services while collaborating with ML engineers, Forward Deployed teams, and product teams.

Qualifications

  • Bachelor's degree in CS or related field.
  • 5+ years of backend/ML infrastructure experience.
  • Strong API design and distributed systems knowledge.
  • Experience with ML/model integration in production.
  • Proficiency with cloud platforms (AWS/GCP/Azure) and containerization.

Responsibilities

  • Design, build, and maintain backend services and APIs powering GenAI, LLM, and CV model integrations.
  • Build and own core AI/ML infrastructure: model-serving pipelines, inference services, data pipelines, and embedding/vector stores.
  • Architect scalable, production-grade systems for real-time and batch AI workloads in manufacturing, infrastructure, and logistics.
  • Implement and optimize RAG systems, prompt/context pipelines, and orchestration layers for enterprise data connections.
  • Develop robust APIs and microservices connecting AI systems to customer data and legacy systems.
  • Ensure reliability, performance, and observability with logging, monitoring, and CI/CD for ML services.
  • Collaborate with Forward Deployed Engineers, ML engineers, and product teams to translate requirements into reusable backend capabilities.
  • Evaluate and integrate ML/CV/LLM models into production backends; manage versioning and deployment pipelines.

Skills

Backend eng
API design
Distributed systems
TypeScript/Node.js
Python ML

Education

Bachelor's degree in CS or related

Tools

Docker
Kubernetes
Cloud platforms
Kafka
gRPC
WebSockets

Job description

Nexxa is seeking a Backend AI Engineer to design, build, and own core AI infrastructure powering production GenAI, ML, and CV solutions. You will create model-serving pipelines, inference services, data pipelines, and APIs to connect models with enterprise data sources.

You’ll architect scalable, real-time and batch systems, implement RAG pipelines, and drive reliability, observability, and CI/CD for ML services while collaborating with ML engineers, Forward Deployed teams, and product teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Backend AI Engineer — Scalable ML Infra & APIs
Backend AI Engineer — Scalable ML Infra & APIs

Nexxa.ai • Sunnyvale (CA)

On-site
USD 140,000 - 220,000
Backend AI Engineer — Scalable ML Infra & APIs
Backend AI Engineer — Scalable ML Infra & APIs

Nexxa.AI • San Francisco (CA)

On-site
USD 140,000 - 230,000
Backend AI Engineer
Backend AI Engineer

Engg • San Francisco (CA)

On-site
USD 180,000 - 240,000
Backend AI Engineer
Backend AI Engineer

Nexxa.AI • San Francisco (CA)

On-site
USD 140,000 - 230,000
Backend AI Engineer
Backend AI Engineer

Nexxa.ai • Sunnyvale (CA)

On-site
USD 140,000 - 220,000
Staff DevOps Engineer - AI Infra for Production at Scale
Staff DevOps Engineer - AI Infra for Production at Scale

Nexxa.AI • San Francisco (CA)

On-site
USD 170,000 - 210,000
Competitive compensation
Equity
Staff DevOps Engineer - AI/ML Infra, Kubernetes & Scale
Staff DevOps Engineer - AI/ML Infra, Kubernetes & Scale

Engg • San Francisco (CA)

On-site
USD 170,000 - 210,000
Staff DevOps Engineer: Industrial AI Infra at Scale
Staff DevOps Engineer: Industrial AI Infra at Scale

Nexxa.ai • Sunnyvale (CA)

On-site
USD 200,000 - 260,000
Equity
Competitive compensation
Remote AI/ML Platform Engineer — Scale Production & Models
Remote AI/ML Platform Engineer — Scale Production & Models

Whatnot • San Francisco (CA)

Hybrid
USD 245,000 - 345,000
Health Insurance options
Work From Home Support
Home office setup allowance
+5
AI Backend Engineer: Build Scalable Inference Pipelines
AI Backend Engineer: Build Scalable Inference Pipelines

NTIATIVE IT Recruitment • Town of Poland (NY)

On-site
USD 120,000 - 190,000
Autonomy from day one
Work with AI technologies
High ownership and impact
+1