Backend Engineer, Multimodal AI Voice Systems

Perplexity

California (MO)

On-site

USD 120,000 - 170,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Perplexity is hiring a backend engineer for the Multimodal team to design and scale live voice sessions and the surrounding infrastructure. You will own end-to-end backend systems, including per-session workers, routing, and real-time streaming.

You will work across the stack with SDK and client teams to ship new voice and multimodal capabilities, build scalable services, and ensure reliability in production environments.

Qualifications

  • 4+ years of professional software engineering experience building backend or distributed systems.
  • Strong experience in Rust, Python, or Go (we work primarily in Rust and Python).
  • Experience designing and operating production distributed systems: streaming RPC (gRPC or similar), stateful services, message-driven architectures, and failure recovery.
  • Solid understanding of cloud infrastructure — deploying, scaling, and operating services on AWS or equivalent.
  • Strong product judgment and the ability to translate user problems into simple, effective technical solutions.
  • Genuine interest and adoption of AI products and willingness to learn quickly.

Responsibilities

  • Design, build, and scale the backend session-worker architecture that powers realtime voice: durable per-session workers, provider routing, and stateful streaming over gRPC.
  • Own distributed-systems problems end-to-end — session lifecycle, crash recovery, reconnection and replay, multi-region deployment, and graceful degradation under real production load.
  • Build provider-agnostic streaming protocols from our backend to the Rust SDK that powers voice across every client stack.
  • Drive new products and initiatives in voice and multimodal AI from problem definition through technical design, implementation, and launch.
  • Build the orchestration layer that lets live voice models delegate work to tools, agents, and long-running tasks — safely, asynchronously, and at scale.
  • Partner closely with SDK, client, infrastructure, and model teams; work across the stack when the product demands it, from backend services to client-facing APIs.

Skills

4+ years backend engineering
Rust
Python
Go

Job description

Perplexity is hiring a backend engineer for the Multimodal team to design and scale live voice sessions and the surrounding infrastructure. You will own end-to-end backend systems, including per-session workers, routing, and real-time streaming.

You will work across the stack with SDK and client teams to ship new voice and multimodal capabilities, build scalable services, and ensure reliability in production environments.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff (AI Software Engineer, Multimodal)
Member of Technical Staff (AI Software Engineer, Multimodal)

Perplexity AI Inc. • San Francisco (CA)

On-site
USD 120,000 - 150,000
Member of Technical Staff (Software Engineer, Multimodal)
Member of Technical Staff (Software Engineer, Multimodal)

Perplexity • California (MO)

On-site
USD 120,000 - 170,000
Multimodal AI Systems Engineer - End-to-End Impact
Multimodal AI Systems Engineer - End-to-End Impact

Perplexity AI Inc. • San Francisco (CA)

On-site
USD 120,000 - 150,000
Staff Engineer: AI Product Systems & Platforms
Staff Engineer: AI Product Systems & Platforms

Perplexity • California (MO)

On-site
USD 120,000 - 190,000
Staff Engineer, Integrations: Agentic AI Apps
Staff Engineer, Integrations: Agentic AI Apps

Perplexity • California (MO)

On-site
USD 140,000 - 210,000
Senior Backend Engineer, Real-Time Voice & AI
Senior Backend Engineer, Real-Time Voice & AI

Fuel Talent LLC • Seattle (WA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Backend Engineer, AI-Driven Voice Systems
Backend Engineer, AI-Driven Voice Systems

Deepgram, Inc. • San Francisco (CA)

On-site
USD 140,000 - 190,000
Multimodal API Backend Engineer
Multimodal API Backend Engineer

OpenAI • California (MO)

On-site
USD 190,000 - 300,000
Staff Software Engineer: AI Acceleration & Platforms
Staff Software Engineer: AI Acceleration & Platforms

Perplexity • California (MO)

Hybrid
USD 180,000 - 280,000
Staff AI Engineer, Agents & Frontier AI
Staff AI Engineer, Agents & Frontier AI

Neura Market • San Francisco (CA)

On-site
USD 140,000 - 190,000