Backend AI Engineer: Low-Latency Inference & Orchestration

United States Digital Space LLC

United States

Remote

USD 120,000 - 150,000

Full time

7 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

ActAI is seeking a Backend Engineer, AI to own the inference and orchestration layer powering AI interactions across the product. You will build and operate production systems turning model capability into fast, stable APIs used by mobile and desktop clients.

Ideal candidates will have strong backend fundamentals, experience with high-throughput services, and familiarity with AI inference patterns, including LLMs and embeddings.

Qualifications

  • Strong backend engineering fundamentals in production environments.
  • Experience running high-throughput, low-latency services.
  • Familiarity with AI inference patterns (LLMs, embeddings, multimodal).
  • Comfortable debugging distributed systems under load.

Responsibilities

  • Build and operate backend systems that serve AI-powered features in production.
  • Design inference pipelines, orchestration layers, and service boundaries around models.
  • Own production concerns: monitoring, logging, alerting, and incident response.
  • Optimize latency and throughput across inference, caching, batching, and streaming.

Skills

Backend engineering
High-throughput services
Debugging distributed systems

Tools

Python
NodeJs
Pytorch
OpenAI/Anthropic LLMs
SQL/NoSQL
Kubernetes
Docker

Job description

ActAI is seeking a Backend Engineer, AI to own the inference and orchestration layer powering AI interactions across the product. You will build and operate production systems turning model capability into fast, stable APIs used by mobile and desktop clients.

Ideal candidates will have strong backend fundamentals, experience with high-throughput services, and familiarity with AI inference patterns, including LLMs and embeddings.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Backend AI Engineer: Inference & Orchestration
Backend AI Engineer: Inference & Orchestration

ActAI • United States

On-site
USD 140,000 - 210,000
Backend AI Engineer: Scalable Inference & Orchestration
Backend AI Engineer: Scalable Inference & Orchestration

re-zoo-me • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
AI Backend Engineer: Inference & Orchestration
AI Backend Engineer: Inference & Orchestration

A1 • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Back End Engineer, AI Systems
Back End Engineer, AI Systems

Salt Digital Recruitment • United States

On-site
USD 140,000 - 190,000
AI Systems Full-Stack Engineer: Orchestrate Reliable Workflows
AI Systems Full-Stack Engineer: Orchestrate Reliable Workflows

Embedded Shishya • United States

On-site
USD 130,000 - 180,000
Backend Engineer, AI (Agent Systems)
Backend Engineer, AI (Agent Systems)

re-zoo-me • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Backend Engineer, Scalable AI Inference & APIs
Backend Engineer, Scalable AI Inference & APIs

Salt Digital Recruitment • United States

On-site
USD 140,000 - 190,000
Senior AI Engineer: Real-Time Inference & Agent Systems
Senior AI Engineer: Real-Time Inference & Agent Systems

Arcana Analytics • United States

On-site
USD 120,000 - 160,000
AI Backend Engineer: Low-Latency Inference Systems
AI Backend Engineer: Low-Latency Inference Systems

A1 • Palo Alto (CA)

On-site
USD 255,000 - 405,000
Backend Engineer, AI Systems
Backend Engineer, AI Systems

A1 • Palo Alto (CA)

On-site
USD 150,000 - 210,000