Founding Inference Engineer, Apple Silicon Platform

Mount Thor

San Francisco (CA)

On-site

USD 240,000 - 320,000

Full time

7 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Mount Thor is hiring a founding inference engineer to build a production inference platform on Apple Silicon. You will own the stack from GPU kernels and model execution to distributed scheduling, networking, and customer-facing serving systems.

You will define the architecture, implement performance-critical code, and shepherd the system from deployment to production operation, ensuring low latency, high throughput, and reliable service for customers.

Qualifications

  • Production inference engines experience and ownership of performance-critical code.
  • Strong systems programming skills in C++ or Rust with Python proficiency.
  • Practical understanding of transformer inference, attention, and quantization.
  • Experience with GPU programming, memory hierarchies, and performance analysis.
  • Strong distributed-systems fundamentals including scheduling and networking.
  • Experience bringing software from architecture to production deployment.

Responsibilities

  • Build and optimize the inference stack from kernels to serving APIs.
  • Design for Apple Silicon, profiling on real hardware and optimizing memory usage.
  • Build distributed inference and its communication layer across machines.
  • Own inference scheduling, routing, and autoscaling to meet latency targets.
  • Ship a production inference service with observability and safe releases.

Skills

C++
Rust
Python
Systems programming
Distributed systems

Tools

MLX
llama.cpp
vLLM
TensorRT-LLM

Job description

Mount Thor is hiring a founding inference engineer to build a production inference platform on Apple Silicon. You will own the stack from GPU kernels and model execution to distributed scheduling, networking, and customer-facing serving systems.

You will define the architecture, implement performance-critical code, and shepherd the system from deployment to production operation, ensuring low latency, high throughput, and reliable service for customers.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Member of Technical Staff, Inference
Member of Technical Staff, Inference

Mount Thor • San Francisco (CA)

On-site
USD 240,000 - 320,000
Developer Relations Lead, macOS & Apple Silicon
Developer Relations Lead, macOS & Apple Silicon

Mount Thor • San Francisco (CA)

On-site
USD 180,000 - 260,000
Senior ML Engineer, Foundation Model Inference (Cloud OS)
Senior ML Engineer, Foundation Model Inference (Cloud OS)

Apple Inc. • Seattle (WA)

On-site
USD 185,000 - 325,000
Senior ML Engineer, Foundation Models Inference — Cloud OS
Senior ML Engineer, Foundation Models Inference — Cloud OS

Apple Inc. • Santa Clara (CA), Northern (KY)

Hybrid
USD 185,000 - 325,000
Medical and dental coverage
Retirement benefits
Employee stock programs
Member of Technical Staff, Platform
Member of Technical Staff, Platform

Mount Thor • San Francisco (CA)

On-site
USD 140,000 - 200,000
Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference
Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference

Apple Inc. • Santa Clara (CA), Northern (KY)

On-site
USD 185,000 - 325,000
Medical and dental coverage
Retirement benefits
Employee stock programs
Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference
Sr. Machine Learning Engineer, Foundation Models Inference - Cloud OS & Inference

Apple Inc. • Seattle (WA)

On-site
USD 185,000 - 325,000
Founding Platform Engineer
Founding Platform Engineer

General Compute • San Francisco (CA)

On-site
USD 180,000 - 260,000
Developer Relations Lead
Developer Relations Lead

Mount Thor • San Francisco (CA)

On-site
USD 180,000 - 260,000
Founding Inference Engineer
Founding Inference Engineer

General Compute • San Francisco (CA)

On-site
USD 180,000 - 320,000