Member of Technical Staff, Infrastructure Engineer

Odyssey

Palo Alto (CA)

On-site

USD 180,000 - 250,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

A pioneering AI lab is seeking an engineer to develop and support low-latency model inference platforms. You will engineer and scale data processing infrastructure while optimizing performance and cost. Ideal candidates should have strong programming skills, experience with container orchestration, and a passion for building efficient systems in a collaborative environment. This role offers the opportunity to shape the future of AI and its applications.

Qualifications

  • Strong programming skills in Python, Go, or similar.
  • Deep experience with containerization and orchestration.
  • Proven experience with GPU computational workloads.

Responsibilities

  • Develop and operate low-latency model inference platforms.
  • Engineer and scale core data processing infrastructure.
  • Design and maintain GPU-based training clusters.
  • Automate infrastructure provisioning and monitoring.
  • Drive performance tuning and cost optimization.
  • Collaborate with researchers and product teams to optimize workflows and usability.

Skills

Programming skills (Python, Go)
Containerization (Docker)
Container orchestration (Kubernetes)
Infrastructure as Code (Terraform)
Experience with distributed systems
Performance tuning
Collaboration and communication skills

Tools

Flyte
Ray
Kubernetes

Job description

Who we are

Odyssey is an AI lab pioneering general‑purpose world models—a new form of multimodal intelligence unlocking entirely new consumer, enterprise, and intelligence applications. World models are the next major frontier in AI, and Odyssey is leading the way with breakthrough models like Odyssey-2 Pro.

What we’re looking for

We are looking for an engineer who thrives on building the engines that make groundbreaking research and products possible. You think in systems, love performance, and get energy from turning theoretical bottlenecks into beautifully efficient reality. You’re excited to design and support infrastructure not just for scale, but for speed, creativity, and discovery. You want to build the compute substrate that lets Odyssey’s world models imagine, act, and interact in real time.

What you’ll do
  • Develop and operate our low‑latency model inference platform, ensuring high availability, scalability, and efficient resource utilization for Odyssey’s world models.
  • Engineer and scale our core data processing infrastructure (e.g., Flyte, Ray with k8s) to handle petabyte‑scale datasets.
  • Design, build, and maintain our large‑scale, GPU‑based training clusters for deep learning, focusing on usability, high throughput and reliability.
  • Automate infrastructure provisioning, configuration, monitoring, and alerting using Infrastructure as Code (IaC) principles.
  • Drive performance tuning, cost optimization, and reliability improvements across the entire stack.
  • Collaborate closely with researchers and product developers to understand their requirements, optimize their workflows, and improve platform usability.
Who you are
  • Motivated by building for the frontier: you want to shape the compute and infrastructure foundation of a lab redefining how people create and interact with media.
  • Strong programming skills (e.g., Python, Go, or similar) and a solid understanding of software engineering best practices.
  • Deep, hands‑on experience with containerization (e.g., Docker), container orchestration (Kubernetes) and Infrastructure as Code (Terraform).
  • Proven experience building and managing large‑scale, distributed systems with GPU computational workloads (e.g., compute platforms, data pipelines, or high‑availability services).
  • Experienced in designing infrastructure for ML workloads where performance, parallelism, and data movement are critical.
  • A collaborative mindset and excellent communication skills, with a passion for building developer‑friendly platforms.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Member of Technical Staff, Core Model Engineering
Member of Technical Staff, Core Model Engineering

Odyssey • Palo Alto (CA), Northern (KY)

Hybrid
USD 180,000 - 320,000
Member of Technical Staff, ML Performance
Member of Technical Staff, ML Performance

Odyssey • Palo Alto (CA)

On-site
USD 130,000 - 160,000
Member of Technical Staff, ML Performance
Member of Technical Staff, ML Performance

Odyssey • Santa Clara (CA)

On-site
USD 120,000 - 160,000
Member of Technical Staff, Data Engineering
Member of Technical Staff, Data Engineering

Odyssey • Palo Alto (CA)

On-site
USD 140,000 - 190,000
Member of Technical Staff, Applied Research
Member of Technical Staff, Applied Research

Odyssey • Palo Alto (CA)

On-site
USD 180,000 - 260,000
Member of Technical Staff, Foundation Models
Member of Technical Staff, Foundation Models

Odyssey • Palo Alto (CA)

On-site
USD 180,000 - 240,000
Member of Technical Staff, TLM: Research Scientist
Member of Technical Staff, TLM: Research Scientist

Odyssey • Palo Alto (CA)

On-site
USD 180,000 - 260,000
VP, Engineering
VP, Engineering

Odyssey • Palo Alto (CA)

On-site
USD 250,000 - 420,000
Member of Technical Staff, Data Engineering
Member of Technical Staff, Data Engineering

Odyssey • Palo Alto (CA)

On-site
USD 120,000 - 180,000
Member of Technical Staff
Member of Technical Staff

Harrison Clarke • San Francisco (CA)

On-site
USD 180,000 - 280,000