Senior AI Runtime Engineer - High-Performance Platform

Socket.dev

United States

Hybrid

USD 216,000 - 324,000

Full time

3 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

RSU grants
Comprehensive healthcare
Paid time off
Team onsite events

Job summary

Modular is building an AI platform to enable developers to deploy state-of-the-art models across hardware. As an AI Runtime Engineer, you will own a runtime that runs on CPU and GPU platforms, optimizing performance for diverse customer models.

Locations: candidates in the US or Canada may apply. You can work in our Los Altos, CA office or remotely from home; onboarding for new hires is conducted in person in Los Altos.

Qualifications

  • 5+ years of experience working on high-performance computing systems.
  • Experience in C++ programming and complex software systems.
  • Experience with CPU or GPU runtime optimizations and performance analysis.
  • Proficiency with profiling tools (CPU or GPU).
  • Creativity and curiosity for solving complex problems with a team-oriented mindset.

Responsibilities

  • Design and develop runtime and cross-stack optimizations to improve CPU and GPU efficiency.
  • Port the Modular runtime stack to new hardware platforms and develop an API.
  • Collaborate with compiler, kernels, serving, and models teams to achieve end-to-end performance.
  • Collaborate with customer success to understand performance requirements and use cases.
  • Collaborate with tooling and infrastructure teams to design automated performance analysis and benchmarking.

Skills

5+ years HPC
C++ programming
CPU/GPU runtime tuning
Profiling tools
Team collaboration

Job description

Modular is building an AI platform to enable developers to deploy state-of-the-art models across hardware. As an AI Runtime Engineer, you will own a runtime that runs on CPU and GPU platforms, optimizing performance for diverse customer models.

Locations: candidates in the US or Canada may apply. You can work in our Los Altos, CA office or remotely from home; onboarding for new hires is conducted in person in Los Altos.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Runtime Engineer
Senior AI Runtime Engineer

Socket.dev • United States

Hybrid
USD 216,000 - 324,000
RSU grants
Comprehensive healthcare
Paid time off
+1
Runtime Engineer: High-Performance AI Compute
Runtime Engineer: High-Performance AI Compute

SambaNova • Palo Alto (CA)

On-site
USD 110,000 - 150,000
95% premium coverage for employee medical insurance
Dental and Vision insurance
Health Savings Account with employer contribution
+2
Senior AI Kernel Engineer — Remote/Hybrid Inference
Senior AI Kernel Engineer — Remote/Hybrid Inference

Modular • United States

Hybrid
USD 198,000 - 286,000
Amazing Team
World-class Benefits
Competitive Compensation
+1
Senior AI Kernel Engineer
Senior AI Kernel Engineer

Modular • United States

Hybrid
USD 198,000 - 286,000
Amazing Team
World-class Benefits
Competitive Compensation
+1
Cloud AI Platform Product Lead (Remote-friendly)
Cloud AI Platform Product Lead (Remote-friendly)

Modular Mailing Systems, Inc. • Los Altos (CA)

Hybrid
USD 223,000 - 345,000
Amazing Team
World-class Benefits
Competitive Compensation
+1
Senior AI Runtime Engineer: Distributed Training & Scale
Senior AI Runtime Engineer: Distributed Training & Scale

FlexAI • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Competitive salary and benefits package
Work on cutting-edge AI infrastructure
High ownership and fast execution
AI/ML Engineer: Next‑Gen Platforms & GPU Workloads
AI/ML Engineer: Next‑Gen Platforms & GPU Workloads

VeeAR Projects Inc. • Sunnyvale (CA)

On-site
USD 140,000 - 210,000
Software Engineer, Model Runtime
Software Engineer, Model Runtime

AI Chopping Block • San Francisco (CA), Northern (KY)

Hybrid
USD 250,000 - 360,000
Senior Multi-Target Runtime Engineer for AI Stack
Senior Multi-Target Runtime Engineer for AI Stack

Lemurian Labs Inc. • Santa Clara (CA)

On-site
USD 180,000 - 240,000
Equity
Medical/Dental/Vision
Retirement savings plan
+1
Performance Modeling Engineer
Performance Modeling Engineer

OpenAI • Los Angeles (CA)

Hybrid
USD 100,000 - 150,000
Relocation assistance
Hybrid work model