Lead Edge ML Engineer – Audio & Embedded Inference

TetraMem - Accelerate The World

San Jose (CA)

On-site

USD 200,000 - 280,000

Full time

2 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

TetraMem - Accelerate The World is hiring a senior ML engineer in San Jose, CA to develop, optimize, and deploy lightweight edge AI models for audio processing on embedded platforms, including FPGA and ASIC solutions.

You will lead ML initiatives, collaborate with hardware and software teams, and explore state-of-the-art techniques to improve efficiency, latency, and power consumption for embedded AI applications, while driving model compression techniques and contributing to open-source

Qualifications

  • 5+ years of relevant industry experience (or a PhD) in Computer Science, Electrical Engineering, Machine Learning, or related fields.
  • Must have prior experience managing a team, serving in a Team Lead role, or demonstrating strong technical leadership and cross-functional coordination capabilities.
  • Strong hands-on experience in machine learning, with a focus on edge AI, on-device inference, and deploying lightweight models on resource-constrained devices.
  • Proficiency in Python and C/C++, with practical experience in ML model optimization and production deployment.
  • Deep experience with model quantization (PTQ/QAT), pruning, knowledge distillation, sparsity, and other compression techniques for efficient edge inference.
  • Hands-on experience developing for or integrating with AI chip SDKs, neural accelerators (NPUs/DSPs), or hardware-specific toolchains (e.g., NVIDIA TensorRT, Qualcomm Neural Processing SDK, ARM Ethos, or similar).
  • Familiarity with edge inference runtimes (ONNX Runtime, ExecuTorch, TVM) and optimizing models for hardware constraints (latency, memory footprint, power consumption).

Responsibilities

  • Develop, optimize, and deploy lightweight machine learning models for edge AI applications, particularly for audio processing.
  • Implement and optimize ML models on embedded platforms, including FPGA and custom ASIC solutions.
  • Work closely with hardware and software teams to integrate ML models into production systems.
  • Research and implement state-of-the-art ML techniques to enhance model efficiency, latency, and power consumption for embedded AI applications.
  • Improve inference efficiency and model compression techniques, including quantization, pruning, and knowledge distillation.
  • Collaborate with cross-functional teams to drive innovation and contribute to the overall system architecture.
  • Provide technical leadership and mentorship to junior engineers.
  • Publish research findings, present at conferences, and contribute to open-source projects when applicable.

Skills

Edge AI
On-device inference
Python
C/C++
Model optimization
Team leadership
Cross-functional collaboration
Knowledge distillation
Quantization
Pruning

Education

PhD

Tools

PyTorch
TensorFlow
TensorFlow Lite
JAX
ONNX Runtime
TVM
TensorRT
NVIDIA TensorRT
Qualcomm Neural Processing SDK
ARM Ethos

Job description

TetraMem - Accelerate The World is hiring a senior ML engineer in San Jose, CA to develop, optimize, and deploy lightweight edge AI models for audio processing on embedded platforms, including FPGA and ASIC solutions.

You will lead ML initiatives, collaborate with hardware and software teams, and explore state-of-the-art techniques to improve efficiency, latency, and power consumption for embedded AI applications, while driving model compression techniques and contributing to open-source

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Edge AI Engineer: Real-Time Audio ML on Embedded Hardware
Edge AI Engineer: Real-Time Audio ML on Embedded Hardware

Bose Corporation • Framingham (MA)

On-site
USD 141,000 - 194,000
Bonus programs
Health & welfare benefits
401(k) plan
+1
Senior Embedded AI Engineer — Edge Audio DSP
Senior Embedded AI Engineer — Edge Audio DSP

5V Tech • San Francisco (CA)

On-site
USD 170,000 - 240,000
Performance bonus
Stock options
Medical, Dental, Vision
+2
Senior Machine Learning Engineer
Senior Machine Learning Engineer

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 200,000 - 280,000
Senior ML Engineer, Edge Audio & Models
Senior ML Engineer, Edge Audio & Models

Syntiant • Redwood City (CA)

On-site
USD 160,000 - 190,000
100% paid medical coverage
Dental PPO coverage
Vision PPO coverage
+5
Edge AI & Audio ML Engineer (On-Device)
Edge AI & Audio ML Engineer (On-Device)

Bose Corporation, U.S.A • Framingham (MA)

On-site
USD 141,000 - 194,000
Bonus programs
Health and welfare benefits
401(k) plan
+1
Edge AI Engineer — On-Device Audio & DSP Innovator
Edge AI Engineer — On-Device Audio & DSP Innovator

Bose • Framingham (MA)

On-site
USD 141,000 - 194,000
Head of Machine Learning
Head of Machine Learning

5V Tech • San Francisco (CA)

On-site
USD 180,000 - 260,000
Equity
Bonus
Medical
+3
Senior Embedded Software Engineer
Senior Embedded Software Engineer

5V Tech • San Francisco (CA)

On-site
USD 170,000 - 240,000
Performance bonus
Stock options
Medical, Dental, Vision
+2
Senior ML Engineer, Edge Audio & Models
Senior ML Engineer, Edge Audio & Models

Syntiant Corp. • Redwood City (CA)

On-site
USD 160,000 - 190,000
Medical coverage fully paid for employees and families
Dental and vision insurance
401k Retirement Plan
+2
Senior Embedded ML Engineer - Real-Time Edge AI
Senior Embedded ML Engineer - Real-Time Edge AI

Allen Control Systems • Austin (TX)

On-site
USD 140,000 - 190,000
Competitive salary
ACS Equity Package
Health, Dental, Vision Insurance
+1