Backend AI Engineer

Nexxa.AI

United States

Remote

USD 120,000 - 180,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Nexxa.AI is seeking a Backend AI Engineer to design, build, and own the core AI infrastructure powering our products in production.

This role blends backend software engineering, ML infrastructure, and systems architecture to design distributed systems, serve models, and deploy reusable backend capabilities across manufacturing, infrastructure, and logistics contexts.

Qualifications

  • 4-8+ years of backend software engineering or ML/platform engineering.
  • Proficiency in TypeScript/Node.js with strong API/microservice design skills; Python a plus for ML tasks.
  • Experience building and operating production backend systems at scale (distributed systems, databases, queues).
  • Experience integrating ML/Generative AI models into backend services (inference, orchestration, evaluation).
  • Knowledge of cloud infrastructure and containerization (Docker/Kubernetes).
  • Experience designing data pipelines (batch/streaming) and building RAG systems.

Responsibilities

  • Design, build, and maintain backend services and APIs powering model integrations.
  • Build and own AI/ML infrastructure: model-serving pipelines, inference services, data pipelines.
  • Architect scalable, production-grade systems for real-time and batch AI workloads.
  • Implement and optimize RAG systems, prompt pipelines, and orchestration layers.
  • Build robust APIs, microservices, and integration layers with enterprise data sources.
  • Own reliability, performance, and observability of backend AI systems.

Skills

TypeScript/Node.js
API design
Distributed systems
Python (ML integration)
Cloud infrastructure
ML/inference integration

Education

Bachelor's degree in Computer Science or related field

Tools

Docker
Kubernetes
AWS

Job description

Nexxa is building the best AI systems for heavy industries - enabling machines, systems and operations to think, decide and act autonomously across manufacturing, large-scale infrastructure, logistics and legacy environments.
Our mission is to translate deep technical breakthroughs into operational reality, solving some of the hardest systems-level problems in industry.

Role Overview

We're looking for Backend AI Engineers to design, build, and own the core AI infrastructure and services that power Nexxa's products in production. Where our Forward Deployed Engineers embed with customers to deliver solutions on the ground, this role builds the systems that make those solutions possible at scale: model-serving pipelines, inference and orchestration layers, data pipelines, and the APIs and microservices that connect Generative AI, Computer Vision, and Machine Learning models to real enterprise environments.

This role is a blend of backend software engineering, ML infrastructure, and systems architecture. You'll design distributed systems, integrate and serve models in production, and build the reusable platform capabilities that our customer-facing and product teams depend on.

Key Responsibilities
  • Design, build, and maintain backend services and APIs that power GenAI, LLM, and Computer Vision model integrations across Nexxa's products.
  • Build and own core AI/ML infrastructure: model-serving pipelines, inference services, data pipelines, and embedding/vector stores.
  • Architect scalable, production-grade systems for real-time and batch AI workloads across manufacturing, infrastructure, and logistics domains.
  • Implement and optimize RAG systems, prompt/context pipelines, and orchestration layers connecting models to enterprise and operational data sources.
  • Build robust APIs, microservices, and integration layers connecting AI systems to customer data, legacy systems, and existing infrastructure.
  • Own the reliability, performance, and observability of backend AI systems - logging, monitoring, testing, and CI/CD for ML services.
  • Collaborate closely with Forward Deployed Engineers, ML engineers, and product teams to translate customer and field requirements into reusable, hardened backend capabilities.
  • Evaluate and integrate ML/CV/LLM models into production backend systems; manage model versioning, rollout, and deployment pipelines.
  • Produce clear technical documentation: architecture diagrams, API specs, and runbooks for internal and customer-facing teams.
  • Mentor engineers and contribute to internal backend engineering best practices.
Qualifications
  • 4-8+ years of experience in backend software engineering, ML/platform engineering, or similar roles.
  • Strong proficiency in TypeScript/Node.js (our primary backend language), with strong API and microservice design skills. Working proficiency in Python is a plus for ML/model integration work.
  • Hands-on experience building and operating production backend systems at scale - distributed systems, databases, message queues.
  • Experience integrating ML or Generative AI models (LLMs, multimodal models) into backend services - inference, orchestration, and evaluation.
  • Solid understanding of cloud infrastructure (AWS, GCP, or Azure) and containerization (Docker, Kubernetes).
  • Experience designing and operating data pipelines (batch and/or streaming) across structured and unstructured data.
  • Hands-on experience building retrieval-augmented generation (RAG) systems and AI memory architectures - retrieval pipelines, vector stores, context management, and long-term/session memory for LLM applications.
  • Strong grasp of system design fundamentals: scalability, reliability, security, and observability.
  • Comfortable working cross-functionally with ML engineers, product, and customer-facing teams.
  • Bachelor's degree (or higher) in Computer Science or a related field.
Preferred
  • Familiarity with ML frameworks (PyTorch, TensorFlow, OpenCV) sufficient to integrate, serve, or evaluate models, even without training them yourself.
  • Experience with MLOps tooling: model registries, feature stores, CI/CD for ML, and monitoring/observability for ML systems.
  • Background in event-driven or real-time systems (Kafka, gRPC, WebSockets).
  • Experience in industrial, IoT, or operational technology (OT) environments.
  • Experience in startup or high-growth environments.
What We’re Looking For
  • A backend engineer who wants to build the infrastructure powering real-world autonomous AI systems.
  • Someone who can architect for scale and reliability while still moving fast.
  • A systems thinker who enjoys turning ambiguous AI capabilities into dependable, production-grade backend services.
  • A strong collaborator who partners well with ML engineers, Forward Deployed teams, and product.
Why Join Nexxa.AI?

Innovative Environment: Build the foundational systems behind groundbreaking AI and automation technologies transforming heavy industries.

Collaborative Culture: Be part of a team that values innovation, discipline, and continuous improvement.

Professional Growth: Benefit from significant opportunities for career development and advancement.

Competitive Compensation: Enjoy a comprehensive salary and equity package reflective of your expertise and contributions.

If you're passionate about backend engineering and eager to build the infrastructure powering advanced AI solutions, we'd love to connect.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Backend AI Engineer
Backend AI Engineer

Nexxa.ai • Sunnyvale (CA)

On-site
USD 140,000 - 220,000
Backend AI Engineer
Backend AI Engineer

Engg • San Francisco (CA)

On-site
USD 180,000 - 240,000
Applied AI Engineer
Applied AI Engineer

nexxa • San Francisco (CA)

On-site
USD 180,000 - 260,000
Staff DevOps Engineer
Staff DevOps Engineer

Nexxa.ai • Sunnyvale (CA)

On-site
USD 200,000 - 260,000
Equity
Competitive compensation
AI Vision Engineer
AI Vision Engineer

nexxa • San Francisco (CA)

On-site
USD 180,000 - 260,000
Competitive compensation
Equity package
Career development
+1
Applied AI Engineer
Applied AI Engineer

Nexxa.ai • Sunnyvale (CA)

On-site
USD 180,000 - 240,000
Innovative Environment
Collaborative Culture
Professional Growth
+1
Staff DevOps Engineer
Staff DevOps Engineer

Engg • San Francisco (CA)

On-site
USD 170,000 - 210,000
Backend AI Engineer — Scalable ML Infra & APIs
Backend AI Engineer — Scalable ML Infra & APIs

Nexxa.ai • Sunnyvale (CA)

On-site
USD 140,000 - 220,000
Software Engineer (Bay Area)
Software Engineer (Bay Area)

Nexla • San Mateo (CA)

Hybrid
USD 150,000 - 180,000
Technical Project Manager
Technical Project Manager

nexxa • San Francisco (CA)

On-site
USD 120,000 - 180,000
Equity
Health benefits
401(k)
+2