ML Runtime Optimization Engineer

Decisive Point

Sunnyvale (CA)

On-site

USD 159,053 - 199,295

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Health insurance
401k retirement benefits
Paid time off
Learning and wellness stipends

Job summary

Decisive Point is seeking a Software Engineer in Sunnyvale, California, with expertise in optimizing machine learning models for embedded systems. This role involves performance optimization for embedded compute platforms, collaborating with ML engineers, and requires strong software development skills.

The ideal candidate has a Bachelor’s degree and at least 3 years of experience in ML accelerators and deep learning frameworks such as PyTorch and JAX. The position offers a competitive salary and comprehensive benefits.

Qualifications

  • 3+ years of experience with ML accelerators, GPU, and SoC architecture.
  • Strong software development skills focused on embedded systems.
  • Experience profiling and optimizing model performance.

Responsibilities

  • Drive ML performance optimization for ADAS/AD stacks.
  • Develop strategies to optimize model inference efficiency.
  • Collaborate on technical efforts with ML engineers.

Skills

Machine Learning optimization
Embedded programming
Deep learning frameworks
Computer Science

Education

Bachelor's in Electrical Engineering or Computer Science

Tools

PyTorch
JAX
ONNX
CUDA
TensorRT

Job description

About the role

We are looking for a software engineer with deep experience in optimizing ML models and deploying them on production‑grade embedded runtime environments. You’ll work across the entire ML framework stack (e.g. PyTorch, JAX, ONNX, TensorRT, CUDA, XLA, Triton).

At Applied Intuition, you will:
  • Drive ML performance optimization on multiple technologies for on‑road and off‑road ADAS / AD stacks targeting deployment on a variety of embedded compute platforms
  • Develop compute usage strategies to optimize efficiency and latency of model inference for compute boards selected by our customers
  • Work on model pruning and quantization, and support deployment on memory constrained platforms
  • Collaborate closely with ML engineers and software developers on technical efforts to find and optimize efficient model architecture solutions
  • Set up methodologies to profile the model performance on target embedded compute platforms and identify performance bottlenecks as part of stack integration
We're looking for someone who has:
  • Bachelors in Electrical Engineering or Computer Science, OR B.Sc. in Computer Science, Mathematics, Physics or a related field
  • 3+ years of experience with ML accelerators, GPU, CPU, SoC architecture and micro‑architecture
  • Strong software development skills with the focus on embedded programming
  • Experience profiling and optimizing model performance on embedded compute platforms
  • Experience in working with deep learning frameworks (e.g., PyTorch, JAX, ONNX, etc.)
Nice to have:
  • M.Sc or PhD in a ML related area
  • Built an ML optimization framework from scratch before
  • Deployed ML solutions to embedded chips for real time robotics applications

Compensation at Applied Intuition for eligible roles includes base salary, equity, and benefits. Base salary is a single component of the total compensation package, which may also include equity in the form of options and/or restricted stock units, comprehensive health, dental, vision, life and disability insurance coverage, 401k retirement benefits with employer match, learning and wellness stipends, and paid time off. Note that benefits are subject to change and may vary based on jurisdiction of employment.

Applied Intuition pay ranges reflect the minimum and maximum intended target base salary for new hire salaries for the position. The actual base salary offered to a successful candidate will additionally be influenced by a variety of factors including experience, credentials & certifications, educational attainment, skill level requirements, interview performance, and the level and scope of the position.

For pay transparency purposes, the base salary range for this full‑time position in the location listed is: $159,053 - $199,295 USD annually.

Applied Intuition is an equal opportunity employer and federal contractor or subcontractor. Consequently, the parties agree that, as applicable, they will abide by the requirements of 41 CFR 60‑1.4(a), 41 CFR 60‑300.5(a) and 41 CFR 60‑741.5(a) and that these regulations prohibit discrimination against qualified individuals based on their status as protected veterans or individuals with disabilities, and prohibit discrimination against all individuals based on race, color, religion, sex, sexual orientation, gender identity or national origin. These regulations require that covered prime contractors and subcontractors take affirmative action to employ and advance in employment individuals without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, protected veteran status or disability. The parties also agree that, as applicable, they will abide by the requirements of Executive Order 13496 (29 CFR Part 471, Appendix A to Subpart A), relating to the notice of employee rights under federal labor laws.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Runtime Optimization Engineer
ML Runtime Optimization Engineer

applied • Sunnyvale (CA)

On-site
USD 159,000 - 200,000
Health, dental, and vision insurance
401k retirement benefits with employer match
Learning and wellness stipends
ML Runtime Optimization Engineer
ML Runtime Optimization Engineer

Applied Intuition Inc. • Sunnyvale (CA)

Hybrid
USD 159,000 - 200,000
Software Engineer - Performance Optimization
Software Engineer - Performance Optimization

applied • Sunnyvale (CA)

On-site
USD 199,000 - 265,000
Equity options
Health insurance
401k retirement benefits
+2
Senior Software Engineer - ML Infrastructure
Senior Software Engineer - ML Infrastructure

Decisive Point • Sunnyvale (CA)

On-site
USD 153,000 - 222,000
Comprehensive health insurance
401k with employer match
Learning and wellness stipends
+1
Software Engineer - Performance Optimization
Software Engineer - Performance Optimization

Applied Intuition • Mountain View (CA)

On-site
USD 199,000 - 265,000
Equity options
Comprehensive health insurance
401(k) retirement benefits with employer match
Embedded AI Engineer – Android Automotive (On-Device Intelligence)
Embedded AI Engineer – Android Automotive (On-Device Intelligence)

Applied Intuition Inc. • Sunnyvale (CA)

On-site
USD 150,000 - 250,000
Equity
Health benefits
401k retirement
Embedded AI Engineer - Android Automotive (On-Device Intelligence)
Embedded AI Engineer - Android Automotive (On-Device Intelligence)

Decisive Point • Sunnyvale (CA)

On-site
USD 150,000 - 250,000
Equity options
401k retirement benefits
Comprehensive health insurance
+1
Senior Software Engineer - ML Training Pipelines
Senior Software Engineer - ML Training Pipelines

Applied Intuition • Sunnyvale (CA)

Hybrid
USD 150,000 - 240,000
Health insurance
401k retirement benefits
Paid time off
+1
Embedded AI Engineer – Android Automotive (On-Device Intelligence)
Embedded AI Engineer – Android Automotive (On-Device Intelligence)

applied • Sunnyvale (CA)

On-site
USD 150,000 - 250,000
Equity options
Comprehensive health benefits
401k retirement with employer match
Senior ML Data Engineer
Senior ML Data Engineer

Socket.dev • Sunnyvale (CA)

On-site
USD 150,000 - 240,000
Health insurance
401k with company match
Paid time off