Embedded ML Inference & Optimization Engineer

Applied Intuition Inc.

Sunnyvale (CA)

Hybrid

USD 159,053 - 199,295

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Applied Intuition Inc. is seeking a software engineer with deep expertise in optimizing ML models for production-grade embedded runtime environments.

You will work across the ML framework stack and target a range of embedded compute platforms used in on- and off-road ADAS/AD stacks. You will collaborate with ML engineers and software developers, profiling model performance, implementing pruning/quantization, and driving efficient architectures for memory-constrained devices.

Qualifications

  • Bachelors in Electrical Engineering or Computer Science, or related field.
  • 3+ years of experience with ML accelerators, GPU/CPU/SoC architecture.
  • Strong software development skills with embedded programming focus.
  • Experience profiling and optimizing model performance on embedded compute platforms.
  • Experience working with deep learning frameworks (PyTorch, JAX, ONNX).

Responsibilities

  • Drive ML performance optimization across technologies for embedded ADAS/AD stacks.
  • Develop compute usage strategies to optimize inference latency and efficiency.
  • Work on model pruning and quantization for memory-constrained platforms.
  • Collaborate with ML engineers and software developers to optimize model architectures.
  • Set up profiling methodologies to identify performance bottlenecks on target hardware.

Skills

ML frameworks experience
Embedded programming
Model profiling
DL frameworks (PyTorch)

Education

Bachelors in Electrical Engineering
Bachelors in Computer Science

Job description

Applied Intuition Inc. is seeking a software engineer with deep expertise in optimizing ML models for production-grade embedded runtime environments.

You will work across the ML framework stack and target a range of embedded compute platforms used in on- and off-road ADAS/AD stacks. You will collaborate with ML engineers and software developers, profiling model performance, implementing pruning/quantization, and driving efficient architectures for memory-constrained devices.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Embedded ML Inference Performance Engineer
Embedded ML Inference Performance Engineer

applied • Sunnyvale (CA)

On-site
USD 159,000 - 200,000
Health, dental, and vision insurance
401k retirement benefits with employer match
Learning and wellness stipends
Embedded ML Inference Optimization Engineer
Embedded ML Inference Optimization Engineer

Decisive Point • Sunnyvale (CA)

On-site
USD 159,000 - 200,000
Health insurance
401k retirement benefits
Paid time off
+1
ML Runtime Optimization Engineer
ML Runtime Optimization Engineer

Decisive Point • Sunnyvale (CA)

On-site
USD 159,000 - 200,000
Health insurance
401k retirement benefits
Paid time off
+1
ML Runtime Optimization Engineer
ML Runtime Optimization Engineer

applied • Sunnyvale (CA)

On-site
USD 159,000 - 200,000
Health, dental, and vision insurance
401k retirement benefits with employer match
Learning and wellness stipends
Senior Embedded Systems Performance Engineer (C++, ML)
Senior Embedded Systems Performance Engineer (C++, ML)

Applied Intuition • Mountain View (CA)

On-site
USD 199,000 - 265,000
Equity options
Comprehensive health insurance
401(k) retirement benefits with employer match
On-Device AI Engineer for Android Automotive
On-Device AI Engineer for Android Automotive

Applied Intuition Inc. • Sunnyvale (CA)

On-site
USD 150,000 - 250,000
Equity
Health benefits
401k retirement
ML Runtime Optimization Engineer
ML Runtime Optimization Engineer

Applied Intuition Inc. • Sunnyvale (CA)

Hybrid
USD 159,000 - 200,000
Senior ML Engineer: AI Inference & Performance Optimizer
Senior ML Engineer: AI Inference & Performance Optimizer

Nebius • Palo Alto (CA)

Hybrid
USD 195,000 - 263,000
Health insurance
401(k) plan
Parental leave
+2
Senior ML Performance Engineer: Scale & Throughput
Senior ML Performance Engineer: Scale & Throughput

NLP PEOPLE • Sunnyvale (CA)

On-site
USD 215,000 - 285,000
Embedded AI Engineer: On-Device ML for Cars
Embedded AI Engineer: On-Device ML for Cars

applied • Sunnyvale (CA)

On-site
USD 150,000 - 250,000
Equity options
Comprehensive health benefits
401k retirement with employer match