Backend AI Engineer: Scalable Inference & Orchestration

re-zoo-me

San Francisco, Northern (CA, KY)

Hybrid

USD 150,000 - 190,000

Full time

6 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

ActAI is seeking a Backend Engineer, AI to own the inference and orchestration layer powering all AI interactions in the product. You will build and operate production systems that turn model capability into fast, stable APIs used across mobile and desktop clients.

You should have strong backend fundamentals, experience with high-throughput, low-latency services, and familiarity with AI inference patterns (LLMs, embeddings, multimodal).

Qualifications

  • Strong backend engineering fundamentals in production environments.
  • Experience running high-throughput, low-latency services.
  • Familiarity with AI inference patterns (LLMs, embeddings, multimodal).
  • Comfortable debugging distributed systems under load.
  • Bias toward shipping and learning from production behaviour.

Responsibilities

  • Build and operate backend systems that serve AI-powered features in production.
  • Design inference pipelines, orchestration layers, and service boundaries around models.
  • Own production concerns: monitoring, logging, alerting, and incident response.
  • Optimize latency and throughput across inference, caching, batching, and streaming.

Job description

ActAI is seeking a Backend Engineer, AI to own the inference and orchestration layer powering all AI interactions in the product. You will build and operate production systems that turn model capability into fast, stable APIs used across mobile and desktop clients.

You should have strong backend fundamentals, experience with high-throughput, low-latency services, and familiarity with AI inference patterns (LLMs, embeddings, multimodal).

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Backend Engineer: Inference & Orchestration
AI Backend Engineer: Inference & Orchestration

A1 • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Back End Engineer, AI Systems
Back End Engineer, AI Systems

Salt Digital Recruitment • United States

On-site
USD 140,000 - 190,000
Backend Engineer, Scalable AI Inference & APIs
Backend Engineer, Scalable AI Inference & APIs

Salt Digital Recruitment • United States

On-site
USD 140,000 - 190,000
AI Backend Engineer: Build Scalable Inference Pipelines
AI Backend Engineer: Build Scalable Inference Pipelines

NTIATIVE IT Recruitment • Town of Poland (NY)

On-site
USD 120,000 - 190,000
Autonomy from day one
Work with AI technologies
High ownership and impact
+1
Backend Engineer, AI (Agent Systems)
Backend Engineer, AI (Agent Systems)

re-zoo-me • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Senior AI Backend Engineer | Equity & Production AI
Senior AI Backend Engineer | Equity & Production AI

Within • San Francisco (CA)

On-site
USD 180,000 - 260,000
Equity
100% Employer-Paid Medical, Dental & 0
Paid Parental Leave
+6
Backend Engineer, AI Systems
Backend Engineer, AI Systems

A1 • Palo Alto (CA)

On-site
USD 150,000 - 210,000
Inference Platform Backend Engineer (Equity & Benefits)
Inference Platform Backend Engineer (Equity & Benefits)

Together • San Francisco (CA)

On-site
USD 160,000 - 250,000
Equity
Health insurance
Competitive compensation
Senior Backend Engineer — AI Agent & Data Orchestration
Senior Backend Engineer — AI Agent & Data Orchestration

People In AI • San Mateo (CA)

Hybrid
USD 250,000 - 300,000
Senior AI Engineer: Real-Time Inference & Agent Systems
Senior AI Engineer: Real-Time Inference & Agent Systems

Arcana Analytics • United States

On-site
USD 120,000 - 160,000