Backend AI Engineer — Scalable ML Infra & APIs

Nexxa.AI

San Francisco (CA)

On-site

USD 140,000 - 230,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Nexxa.AI is building the core backend infrastructure powering Generative AI, CV, and ML-enabled products for manufacturing, infrastructure, and logistics. The Backend AI Engineer will design and own model-serving pipelines, inference layers, and data pipelines at scale.

You’ll implement and optimize RAG workflows, build robust APIs, and ensure observability, security, and reliability across production systems. Strong collaboration with ML engineers and product teams is essential.

Qualifications

  • 4–8+ years of backend/software ML or platform engineering.
  • Strong TypeScript/Node.js with API/microservice design; Python is a plus.
  • Experience building production backend systems at scale (distributed systems, databases, queues).
  • Experience integrating ML/Generative AI models into backend services (inference, orchestration).
  • Solid understanding of cloud infrastructure (AWS/GCP/Azure) and containers (Docker, Kubernetes).
  • Experience designing data pipelines (batch/streaming) across structured/unstructured data.
  • Hands-on experience with retrieval-augmented generation (RAG) and vector stores.
  • Strong system design, reliability, security, and observability.
  • Collaborative with ML engineers, product, and customer teams.
  • Bachelor's degree in Computer Science or related field.

Responsibilities

  • Design, build, and maintain backend services and APIs powering AI/model integrations.
  • Build and own core AI/ML infrastructure: model-serving pipelines, inference services, data pipelines.
  • Architect scalable systems for real-time and batch AI workloads across domains.
  • Implement RAG, prompt/context pipelines, orchestration layers.
  • Build APIs and microservices connecting AI systems to data sources.
  • Ensure reliability, performance, and observability of backend AI systems.
  • Collaborate with Forward Deployed teams and product groups to translate requirements.
  • Evaluate and integrate ML/CV/LLM models into production backend systems.
  • Produce architecture diagrams, API specs, and runbooks.
  • Mentor engineers and contribute to backend engineering practices.

Skills

Years-experience
TypeScript/Node.js
Python
Distributed systems
ML model integration
Cloud infrastructure
Docker/Kubernetes
Data pipelines
RAG systems
System design
Cross-functional

Education

Bachelor's degree in CS or related

Tools

Docker
Kubernetes
Kafka
gRPC/WebSockets

Job description

Nexxa.AI is building the core backend infrastructure powering Generative AI, CV, and ML-enabled products for manufacturing, infrastructure, and logistics. The Backend AI Engineer will design and own model-serving pipelines, inference layers, and data pipelines at scale.

You’ll implement and optimize RAG workflows, build robust APIs, and ensure observability, security, and reliability across production systems. Strong collaboration with ML engineers and product teams is essential.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Backend AI Engineer — Scalable ML Infra & APIs
Backend AI Engineer — Scalable ML Infra & APIs

Nexxa.ai • Sunnyvale (CA)

On-site
USD 140,000 - 220,000
Senior Backend AI Engineer: GenAI & ML Infra
Senior Backend AI Engineer: GenAI & ML Infra

Uncover • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Backend AI Engineer
Backend AI Engineer

Nexxa.AI • San Francisco (CA)

On-site
USD 140,000 - 230,000
Backend AI Engineer
Backend AI Engineer

Uncover • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Backend AI Engineer
Backend AI Engineer

Nexxa.ai • Sunnyvale (CA)

On-site
USD 140,000 - 220,000
Staff DevOps Engineer - AI Infra for Production at Scale
Staff DevOps Engineer - AI Infra for Production at Scale

Nexxa.AI • San Francisco (CA)

On-site
USD 170,000 - 210,000
Competitive compensation
Equity
AI Backend Engineer: Build Scalable Inference Pipelines
AI Backend Engineer: Build Scalable Inference Pipelines

NTIATIVE IT Recruitment • Town of Poland (NY)

On-site
USD 120,000 - 190,000
Autonomy from day one
Work with AI technologies
High ownership and impact
+1
Real-Time Ad Tech ML Engineer — Scalable AI & Optimization
Real-Time Ad Tech ML Engineer — Scalable AI & Optimization

Nexxen • Baltimore (MD)

Hybrid
USD 180,000 - 220,000
Medical
Dental
Vision
+6
GenAI Backend Engineer for Scalable AI Systems
GenAI Backend Engineer for Scalable AI Systems

Next Insurance • Boston (MA)

Hybrid
USD 127,000 - 220,000
Go Backend Engineer - Scale Secure Microservices
Go Backend Engineer - Scale Secure Microservices

nexos.ai • Town of Poland (NY)

On-site
Professional growth opportunities
Private health insurance
Additional vacation days
+1