Backend AI Engineer

Engg

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Nexxa is seeking a Backend AI Engineer to design, build, and own core AI infrastructure powering production GenAI, ML, and CV solutions. You will create model-serving pipelines, inference services, data pipelines, and APIs to connect models with enterprise data sources.

You’ll architect scalable, real-time and batch systems, implement RAG pipelines, and drive reliability, observability, and CI/CD for ML services while collaborating with ML engineers, Forward Deployed teams, and product teams.

Qualifications

  • Bachelor's degree in CS or related field.
  • 5+ years of backend/ML infrastructure experience.
  • Strong API design and distributed systems knowledge.
  • Experience with ML/model integration in production.
  • Proficiency with cloud platforms (AWS/GCP/Azure) and containerization.

Responsibilities

  • Design, build, and maintain backend services and APIs powering GenAI, LLM, and CV model integrations.
  • Build and own core AI/ML infrastructure: model-serving pipelines, inference services, data pipelines, and embedding/vector stores.
  • Architect scalable, production-grade systems for real-time and batch AI workloads in manufacturing, infrastructure, and logistics.
  • Implement and optimize RAG systems, prompt/context pipelines, and orchestration layers for enterprise data connections.
  • Develop robust APIs and microservices connecting AI systems to customer data and legacy systems.
  • Ensure reliability, performance, and observability with logging, monitoring, and CI/CD for ML services.
  • Collaborate with Forward Deployed Engineers, ML engineers, and product teams to translate requirements into reusable backend capabilities.
  • Evaluate and integrate ML/CV/LLM models into production backends; manage versioning and deployment pipelines.

Skills

Backend eng
API design
Distributed systems
TypeScript/Node.js
Python ML

Education

Bachelor's degree in CS or related

Tools

Docker
Kubernetes
Cloud platforms
Kafka
gRPC
WebSockets

Job description

Nexxa http://Nexxa.ai is building the best AI systems for heavy industries — enabling machines, systems and operations to think, decide and act autonomously across manufacturing, large-scale infrastructure, logistics and legacy environments. Our mission is to translate deep technical breakthroughs into operational reality, solving some of the hardest systems-level problems in industry.

ROLE OVERVIEW

We're looking for Backend AI Engineers to design, build, and own the core AI infrastructure and services that power Nexxa's products in production. Where our Forward Deployed Engineers embed with customers to deliver solutions on the ground, this role builds the systems that make those solutions possible at scale: model-serving pipelines, inference and orchestration layers, data pipelines, and the APIs and microservices that connect Generative AI, Computer Vision, and Machine Learning models to real enterprise environments. This role is a blend of backend software engineering, ML infrastructure, and systems architecture. You'll design distributed systems, integrate and serve models in production, and build the reusable platform capabilities that our customer-facing and product teams depend on.

KEY RESPONSIBILITIES
  • Design, build, and maintain backend services and APIs that power GenAI, LLM, and Computer Vision model integrations across Nexxa's products.
  • Build and own core AI/ML infrastructure: model-serving pipelines, inference services, data pipelines, and embedding/vector stores.
  • Architect scalable, production-grade systems for real-time and batch AI workloads across manufacturing, infrastructure, and logistics domains.
  • Implement and optimize RAG systems, prompt/context pipelines, and orchestration layers connecting models to enterprise and operational data sources.
  • Build robust APIs, microservices, and integration layers connecting AI systems to customer data, legacy systems, and existing infrastructure.
  • Own the reliability, performance, and observability of backend AI systems — logging, monitoring, testing, and CI/CD for ML services.
  • Collaborate closely with Forward Deployed Engineers, ML engineers, and product teams to translate customer and field requirements into reusable, hardened backend capabilities.
  • Evaluate and integrate ML/CV/LLM models into production backend systems; manage model versioning, rollout, and deployment pipelines.
  • Produce clear technical documentation: architecture diagrams, API specs, and runbooks for internal and customer-facing teams.
  • Mentor engineers and contribute to internal backend engineering best practices.
QUALIFICATIONS
  • 4-8+ years of experience in backend software engineering, ML/platform engineering, or similar roles.
  • Strong proficiency in TypeScript/Node.js (our primary backend language), with strong API and microservice design skills. Working proficiency in Python is a plus for ML/model integration work.
  • Hands-on experience building and operating production backend systems at scale — distributed systems, databases, message queues.
  • Experience integrating ML or Generative AI models (LLMs, multimodal models) into backend services — inference, orchestration, and evaluation.
  • Solid understanding of cloud infrastructure (AWS, GCP, or Azure) and containerization (Docker, Kubernetes).
  • Experience designing and operating data pipelines (batch and/or streaming) across structured and unstructured data.
  • Hands-on experience building retrieval-augmented generation (RAG) systems and AI memory architectures — retrieval pipelines, vector stores, context management, and long-term/session memory for LLM applications.
  • Strong grasp of system design fundamentals: scalability, reliability, security, and observability.
  • Comfortable working cross-functionally with ML engineers, product, and customer-facing teams.
  • Bachelor's degree (or higher) in Computer Science or a related field.
PREFERRED
  • Familiarity with ML frameworks (PyTorch, TensorFlow, OpenCV) sufficient to integrate, serve, or evaluate models, even without training them yourself.
  • Experience with MLOps tooling: model registries, feature stores, CI/CD for ML, and monitoring/observability for ML systems.
  • Background in event-driven or real-time systems (Kafka, gRPC, WebSockets).
  • Experience in industrial, IoT, or operational technology (OT) environments.
  • Experience in startup or high-growth environments.
WHAT WE'RE LOOKING FOR
  • A backend engineer who wants to build the infrastructure powering real-world autonomous AI systems.
  • Someone who can architect for scale and reliability while still moving fast.
  • A systems thinker who enjoys turning ambiguous AI capabilities into dependable, production-grade backend services.
  • A strong collaborator who partners well with ML engineers, Forward Deployed teams, and product.
WHY JOIN NEXXA.AI
  • Innovative Environment: Build the foundational systems behind groundbreaking AI and automation technologies transforming heavy industries.
  • Collaborative Culture: Be part of a team that values innovation, discipline, and continuous improvement.
  • Professional Growth: Benefit from significant opportunities for career development and advancement.
  • Competitive Compensation: Enjoy a comprehensive salary and equity package reflective of your expertise and contributions.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Backend AI Engineer
Backend AI Engineer

Nexxa.ai • Sunnyvale (CA)

On-site
USD 140,000 - 220,000
Backend AI Engineer
Backend AI Engineer

Nexxa.AI • San Francisco (CA)

On-site
USD 140,000 - 230,000
Staff DevOps Engineer
Staff DevOps Engineer

Nexxa.AI • San Francisco (CA)

On-site
USD 170,000 - 210,000
Competitive compensation
Equity
Staff DevOps Engineer
Staff DevOps Engineer

Engg • San Francisco (CA)

On-site
USD 170,000 - 210,000
AI Vision Engineer
AI Vision Engineer

Nexxa.AI • San Francisco (CA)

On-site
USD 150,000 - 210,000
Competitive salary and equity
Professional growth
Collaborative culture
Staff DevOps Engineer
Staff DevOps Engineer

Nexxa.ai • Sunnyvale (CA)

On-site
USD 200,000 - 260,000
Equity
Competitive compensation
Backend AI Engineer — Scalable ML Infra & APIs
Backend AI Engineer — Scalable ML Infra & APIs

Nexxa.ai • Sunnyvale (CA)

On-site
USD 140,000 - 220,000
Backend AI Engineer — Scalable ML Infra & APIs
Backend AI Engineer — Scalable ML Infra & APIs

Nexxa.AI • San Francisco (CA)

On-site
USD 140,000 - 230,000
QA Engineer (AI Systems)
QA Engineer (AI Systems)

Nexxa.AI • Sunnyvale (CA)

On-site
USD 140,000 - 200,000
Competitive compensation
Equity package
Applied AI Engineer
Applied AI Engineer

Nexxa.ai • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
Innovative Environment
Collaborative Culture
Professional Growth
+1