Runtime Engineer

Lemurian Labs Inc.

Santa Clara (CA)

On-site

USD 180,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity
Medical/Dental/Vision
Retirement savings plan
Wellness benefits

Job summary

Lemurian Labs Inc. in Santa Clara, CA, is hiring a Runtime Engineer to design and build the multi-target runtime at the heart of our AI compiler stack.

This is a systems role focused on low-level parallelization, kernel scheduling, performance analysis, and cross-team collaboration to push AI hardware boundaries. You’ll work with compiler and product teams to optimize execution across diverse targets.

Qualifications

  • BS degree in Computer Science, Computer Engineering, or equivalent practical experience.
  • 4+ years of experience with compilers or runtime systems.
  • Deep understanding of asynchronous and concurrent programming.
  • 4+ years of experience with C/C++ (C++14 or newer).
  • Understanding of hardware architecture and memory hierarchies.
  • Knowledge of OS kernel or hypervisor development.

Responsibilities

  • Design, develop, maintain, and improve our multi-target runtime.
  • Apply parallelization and partitioning techniques to optimize kernel execution paths.
  • Prototype and data-drive exploration of new runtime ideas.
  • Benchmark and analyze compiler outputs on target hardware.
  • Build tools to collect and analyze performance bottlenecks.
  • Collaborate with product teams to meet ML engineer needs and evolve runtime architecture.

Skills

Compiler construction
Runtime systems
Asynchronous programming
C/C++
Parallelization
Kernel scheduling
Performance analysis

Education

BS in CS/CE
Master's/PhD (preferred)

Tools

CUDA
ROCm

Job description

At Lemurian Labs, we're reimagining the foundations of computing to make AI accessible to everyone. Our mission is to remove the limits of scale, hardware, and cost that hold back innovation, so the people solving humanity's hardest problems can move faster.

We're building a new kind of software stack: a hardware-agnostic platform that makes every system — from a laptop to a supercomputer — feel like one seamless engine. Developers can write once, run anywhere, and get state-of-the-art performance across any chip, any cloud, at any scale. It's a complete rethink of how software and hardware interact — designed for the era beyond Moore's Law.

We're not looking for the comfortable or the conventional; we're looking for the bold. The engineers who crave frontier problems, who want to bend the limits of what's possible, who see infrastructure not as a constraint but as a canvas. If you want to build the foundation for the next era of AI and change what humanity can achieve in the process, join us.

About the Role

We're looking for a Runtime Engineer to design and build the multi-target runtime that sits at the heart of our AI compiler stack. This is a systems-level role where you'll take the output of our optimizing compiler and make it execute — efficiently, correctly, and at scale — across a diverse landscape of hardware targets.

You’ll work on low-level parallelization, kernel scheduling, and performance analysis, and collaborate closely with our compiler and product teams to push the boundaries of what's possible on modern AI hardware.

What You’ll Do
  • Design, develop, maintain, and improve our multi-target runtime.
  • Apply the latest techniques in parallelization and partitioning to automate kernel generation and exploit highly optimized execution paths.
  • Rapidly prototype and data-drive exploration of new runtime ideas.
  • Benchmark and analyze the outputs produced by our optimizing compiler on target hardware.
  • Build tools to collect and analyze performance bottlenecks.
  • Work closely with our product team to understand the evolving needs of ML engineers and drive improvements in runtime architecture.
Requirements
Essential Skills and Experience
  • BS degree in Computer Science, Computer Engineering, or equivalent practical experience.
  • 4+ years of experience working with compilers or runtime systems.
  • Deep understanding of asynchronous and concurrent programming.
  • 4+ years of experience with C/C++ (C++14 or newer).
  • Understanding of hardware architecture: vector vs. scalar registers and instructions, memory hierarchies.
  • Knowledge of operating system kernel development or hypervisor development.
Preferred Skills and Experience
  • Master's or PhD in Computer Science, Computer Engineering, or equivalent.
  • Experience developing or maintaining GPU compute libraries such as CUDA or ROCm.
  • Experience with GPU programming and optimization.
  • Background in high-performance computing (HPC).
  • Knowledge of deep learning frameworks such as PyTorch, JAX, or Triton.
  • Experience programming large compute clusters.
Why Join Lemurian Labs
  • Build the runtime that makes next-generation AI infrastructure actually go fast.
  • Work across the full stack — from hardware intrinsics to compiler output to distributed execution.
  • Join a team that approaches infrastructure as a canvas, not a constraint.
  • Competitive compensation including equity, medical/dental/vision, retirement savings, and wellness benefits.

Lemurian Labs is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees, regardless of gender identity, race, ethnicity, sexual orientation, disability status, age, or background.

Compensation depends on experience and geographic location and will be narrowed during the interview process. Additional benefits include equity, company bonus opportunities, medical, dental, and vision coverage, a retirement savings plan, and supplemental wellness benefits.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Compiler Code Gen Engineer
Compiler Code Gen Engineer

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 170,000 - 230,000
Equity
Medical/dental/vision
Retirement savings plan
+1
Runtime Engineer
Runtime Engineer

Amadeus Search • San Francisco (CA)

Hybrid
USD 120,000 - 160,000
Equity opportunities
Medical, dental, and vision coverage
Retirement savings plan
+2
Runtime Engineer
Runtime Engineer

Oho Group • San Francisco (CA)

On-site
USD 180,000 - 240,000
Compiler Optimization Engineer
Compiler Optimization Engineer

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 180,000 - 230,000
Equity
Medical benefits
Retirement plan
+1
Senior DSL Engineer
Senior DSL Engineer

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 170,000 - 250,000
Equity
Medical/Dental/Vision
Retirement savings
+1
Front End Compiler
Front End Compiler

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 140,000 - 190,000
Equity
Medical/Dental/Vision
Retirement savings plan
+1
Senior Developer Tools Engineer
Senior Developer Tools Engineer

Lemurian Labs • Santa Clara (CA)

On-site
USD 100,000 - 130,000
Equity
Medical, dental, and vision coverage
Retirement savings plan
+1
Senior Multi-Target Runtime Engineer for AI Stack
Senior Multi-Target Runtime Engineer for AI Stack

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Equity
Medical/Dental/Vision
Retirement savings plan
+1
Senior Developer Tools Engineer
Senior Developer Tools Engineer

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 140,000 - 190,000
Equity
Medical/dental/vision
Retirement savings
+1
Senior DSL Engineer
Senior DSL Engineer

Lemurian Labs • California (MO)

On-site