Portable AI Runtime & Scheduling Engineer

Oho Group

San Francisco (CA)

On-site

USD 150,000 - 210,000

Full time

8 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Oho Group is seeking a Runtime Engineer to turn optimized compiler output into reliable, high-performance execution for AI workloads. You will design a portable runtime, enabling parallel execution, scheduling, and profiling across diverse hardware.

You will collaborate with compiler and hardware engineers on runtime architecture, profiling, and performance analysis to push the platform forward for real ML workloads and deployment needs.

Qualifications

  • 4+ years in compiler, runtime or low-level systems engineering.
  • Strong modern C++ proficiency.
  • Deep knowledge of concurrency and asynchronous execution.
  • Understanding of memory hierarchies and vector/scalar compute.
  • Experience with OS kernels, hypervisors or similarly low-level systems.

Responsibilities

  • Design a portable runtime for AI workloads.
  • Build workload partitioning, parallel execution and kernel scheduling capabilities.
  • Prototype new runtime techniques and evaluate them against real hardware.
  • Benchmark compiled workloads and diagnose performance bottlenecks.
  • Develop profiling tools that inform compiler and runtime improvements.
  • Help evolve the platform around real ML workloads and deployment requirements.

Skills

Modern C++
Concurrency
Memory hierarchy
Low-level systems

Tools

CUDA
ROCm
PyTorch
JAX
Triton
GPU programming

Job description

Oho Group is seeking a Runtime Engineer to turn optimized compiler output into reliable, high-performance execution for AI workloads. You will design a portable runtime, enabling parallel execution, scheduling, and profiling across diverse hardware.

You will collaborate with compiler and hardware engineers on runtime architecture, profiling, and performance analysis to push the platform forward for real ML workloads and deployment needs.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Compiler Runtime Engineer
Compiler Runtime Engineer

Oho Group • San Francisco (CA)

On-site
USD 150,000 - 210,000
AI Accelerator Runtime Systems Engineer
AI Accelerator Runtime Systems Engineer

Triwill Group • United States

On-site
USD 180,000 - 240,000
Senior AI Compute Architect — GPUs, MLIR & CUDA
Senior AI Compute Architect — GPUs, MLIR & CUDA

Oho Group • San Francisco (CA)

On-site
USD 260,000 - 380,000
Runtime Engineer
Runtime Engineer

Amadeus Search • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Equity opportunities
Medical, dental, and vision coverage
Retirement savings plan
+2
Senior Multi-Target Runtime Engineer for AI Stack
Senior Multi-Target Runtime Engineer for AI Stack

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Equity
Medical/Dental/Vision
Retirement savings plan
+1
Low-Level AI Accelerator Runtime Engineer
Low-Level AI Accelerator Runtime Engineer

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
AI Runtime Engineer
AI Runtime Engineer

Insider, Inc. • United States

On-site
USD 100,000 - 140,000
AI Accelerator Runtime Engineer (Low-Level Systems)
AI Accelerator Runtime Engineer (Low-Level Systems)

OpenAI • San Francisco (CA)

On-site
USD 266,000 - 445,000
Lead AI Compiler & Runtime Systems Engineer
Lead AI Compiler & Runtime Systems Engineer

Xcelerium • Irvine (CA)

Hybrid
USD 190,000 - 260,000
Comprehensive medical, dental, and vision coverage
401(k) plan
Paid time off (PTO)
+1
Runtime Engineer: High-Performance AI Compute
Runtime Engineer: High-Performance AI Compute

SambaNova • Palo Alto (CA)

On-site
USD 110,000 - 150,000
95% premium coverage for employee medical insurance
Dental and Vision insurance
Health Savings Account with employer contribution
+2