Runtime Engineer

Oho Group

San Francisco (CA)

On-site

USD 180,000 - 240,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Oho Group in San Francisco is seeking a Runtime Engineer to build the execution layer of their compiler platform for AI workloads. You will translate optimized compiler output into efficient, scalable code across modern hardware.

This role focuses on parallel execution, kernel scheduling, runtime architecture, and performance analysis, working close to hardware and with compiler engineers to turn ideas into real-world gains.

Qualifications

  • 4+ years working on compiler or runtime systems.
  • Strong modern C++.
  • Deep understanding of asynchronous and concurrent programming.
  • Familiarity with memory hierarchies and vector/scalar execution.
  • OS kernel or hypervisor development experience.
  • GPU programming with CUDA/ROCm, HPC, PyTorch, JAX or Triton a plus.

Responsibilities

  • Design and develop a multi-target runtime for AI workloads.
  • Build parallelisation, partitioning and kernel-scheduling capabilities.
  • Prototype and evaluate new runtime approaches using performance data.
  • Benchmark compiler output on target hardware and diagnose bottlenecks.
  • Develop profiling and analysis tools that guide runtime and compiler improvements.
  • Work with product and compiler teams to evolve the platform around real ML-engineering needs.

Skills

Compiler/runtime experience
Modern C++
Async/Concurrent
Memory Hierarchy
Kernel/Hypervisor
GPU CUDA/ROCm / Triton
PyTorch/JAX/Triton experience

Job description

We are working with an ambitious AI infrastructure company building a new software stack that enables machine-learning workloads to run efficiently across diverse hardware targets.

They are looking for a Runtime Engineer to build the execution layer at the heart of their compiler platform. You will take optimised compiler output and make it run efficiently, correctly and at scale across modern AI hardware.

This is a deeply systems-focused role spanning parallel execution, kernel scheduling, runtime architecture and performance analysis. You will work close to hardware while collaborating directly with compiler engineers to turn optimisation ideas into real-world performance.

What you’ll do

  • Design and develop a multi-target runtime for AI workloads
  • Build parallelisation, partitioning and kernel-scheduling capabilities
  • Prototype and evaluate new runtime approaches using performance data
  • Benchmark compiler output on target hardware and diagnose bottlenecks
  • Develop profiling and analysis tools that guide runtime and compiler improvements
  • Work with product and compiler teams to evolve the platform around real ML-engineering needs

What we’re looking for

  • 4+ years working on compiler or runtime systems
  • Strong modern C++
  • Deep understanding of asynchronous and concurrent programming
  • Familiarity with memory hierarchies and vector/scalar execution
  • OS kernel or hypervisor development experience
  • GPU programming, CUDA/ROCm, HPC, large clusters, PyTorch, JAX or Triton experience would be valuable

You will build the runtime that determines whether next-generation AI infrastructure merely works or genuinely performs. It is a rare opportunity to work across hardware intrinsics, compiler output and distributed execution, with real ownership over a foundational systems product.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Runtime Engineer
Runtime Engineer

Amadeus Search • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Equity opportunities
Medical, dental, and vision coverage
Retirement savings plan
+2
Runtime Engineer
Runtime Engineer

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Equity
Medical/Dental/Vision
Retirement savings plan
+1
AI Runtime Engineer: Multi-Target Performance & Scheduling
AI Runtime Engineer: Multi-Target Performance & Scheduling

Oho Group • San Francisco (CA)

On-site
USD 180,000 - 240,000
AI Runtime Engineer
AI Runtime Engineer

Insider, Inc. • United States

On-site
USD 100,000 - 140,000
Senior Multi-Target Runtime Engineer for AI Stack
Senior Multi-Target Runtime Engineer for AI Stack

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Equity
Medical/Dental/Vision
Retirement savings plan
+1
Runtime Engineer, Multi-Target AI Infrastructure
Runtime Engineer, Multi-Target AI Infrastructure

Amadeus Search • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Equity opportunities
Medical, dental, and vision coverage
Retirement savings plan
+2
Senior Software Engineer – AI Compiler & Runtime Infrastructure
Senior Software Engineer – AI Compiler & Runtime Infrastructure

Xcelerium • Irvine (CA)

Hybrid
USD 190,000 - 260,000
Comprehensive medical, dental, and vision coverage
401(k) plan
Paid time off (PTO)
+1
Senior Software Engineer - AI Compiler & Runtime Infrastructure
Senior Software Engineer - AI Compiler & Runtime Infrastructure

Xcelerium • Irvine (CA)

Hybrid
USD 190,000 - 260,000
Comprehensive medical, dental, and vision coverage
Flexible hybrid work model
401(k) plan
+2
AI Compiler Engineer
AI Compiler Engineer

Darwin Recruitment • San Jose (CA)

On-site
USD 120,000 - 180,000
Runtime SME - AI Infrastructure
Runtime SME - AI Infrastructure

Hamilton Barnes Associates Limited • United States

On-site
USD 212,500 - 287,500
Full Benefits