Senior Machine Learning Engineer

TetraMem - Accelerate The World

San Jose (CA)

On-site

USD 200,000 - 280,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

TetraMem - Accelerate The World seeks a senior ML leader to design and optimize edge AI models for audio processing on embedded platforms. You will guide cross-functional teams, mentor junior engineers, and push for efficient, low-latency implementations across FPGA, ASIC, and software stacks.

You will shape the architecture, champion model compression techniques, and publish or contribute to open-source efforts as part of a broader innovation strategy.

Qualifications

  • 5+ years of industry experience (or a PhD) in CS, EE, ML, or related field.
  • Demonstrated leadership: team lead or cross-functional coordination.
  • Strong hands-on ML with edge AI, on-device inference, and lightweight models on constrained devices.

Responsibilities

  • Develop, optimize, and deploy lightweight ML models for edge AI applications (audio processing).
  • Implement and optimize ML models on embedded platforms (FPGA, ASIC).
  • Collaborate with hardware and software teams to integrate ML into production systems.
  • Research state-of-the-art ML techniques to enhance efficiency, latency, and power on embedded AI.

Skills

Edge AI
On-device inference
Python
C/C++
ML optimization
TensorRT/TFLite

Education

PhD or MSc in CS/EE/ML

Tools

PyTorch
TensorFlow
JAX
ONNX Runtime
NPU/DSP SDKs

Job description

Responsibilities
  • Develop, optimize, and deploy lightweight machine learning models for edge AI applications, particularly for audio processing.
  • Implement and optimize ML models on embedded platforms, including FPGA and custom ASIC solutions.
  • Work closely with hardware and software teams to integrate ML models into production systems.
  • Research and implement state-of-the-art ML techniques to enhance model efficiency, latency, and power consumption for embedded AI applications.
  • Improve inference efficiency and model compression techniques, including quantization, pruning, and knowledge distillation.
  • Collaborate with cross-functional teams to drive innovation and contribute to the overall system architecture.
  • Provide technical leadership and mentorship to junior engineers.
  • Publish research findings, present at conferences, and contribute to open-source projects when applicable.
Requirements
  • 5+ years of relevant industry experience (or a PhD) in Computer Science, Electrical Engineering, Machine Learning, or related fields.
  • Must have prior experience managing a team, serving in a Team Lead role, or demonstrating strong technical leadership and cross-functional coordination capabilities.
  • Strong hands-on experience in machine learning, with a focus on edge AI, on-device inference, and deploying lightweight models on resource-constrained devices.
  • Expertise in modern ML frameworks such as PyTorch, TensorFlow (including TensorFlow Lite), and JAX.
  • Proficiency in Python and C/C++, with practical experience in ML model optimization and production deployment.
  • Deep experience with model quantization (PTQ/QAT), pruning, knowledge distillation, sparsity, and other compression techniques for efficient edge inference.
  • Hands-on experience developing for or integrating with AI chip SDKs, neural accelerators (NPUs/DSPs), or hardware-specific toolchains (e.g., NVIDIA TensorRT, Qualcomm Neural Processing SDK, ARM Ethos, or similar).
  • Familiarity with edge inference runtimes (ONNX Runtime, ExecuTorch, TVM) and optimizing models for hardware constraints (latency, memory footprint, power consumption).
Experience in one or more of the following areas considered a strong plus:
  • Understanding of ML compiler and runtime design.
  • Experience working with tools such as Optimum, ONNX, TensorRT, TFLite/LiteRT, ncnn, or CoreML.
  • Familiarity with hardware acceleration techniques.
  • Experience in embedded system development.
Salary Range:

$200,000 - $280,000 / year

TetraMem celebrates diversity and is committed to creating an inclusive environment for all employees. We are proud to be an Equal Opportunity Employer and welcome applicants from all backgrounds. Qualified candidates will receive consideration for employment without regard to race, color, religion, creed, sex, gender identity or expression, sexual orientation, national origin, ancestry, age, marital status, medical condition, disability, genetic information, military or veteran status, or any other characteristic protected by applicable federal, state, or local law.

TetraMem is committed to providing reasonable accommodations to qualified applicants with disabilities throughout the recruitment process. Applicants requiring accommodation may contact Human Resources for assistance.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

System Engineer
System Engineer

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 110,000 - 250,000
Neuronetwork System Engineer
Neuronetwork System Engineer

TetraMem Inc. • San Jose (CA)

Hybrid
USD 110,000 - 250,000
System Engineer
System Engineer

TETRAMEM INC • San Jose (CA)

On-site
USD 110,000 - 250,000
Neuronetwork System Engineer
Neuronetwork System Engineer

TetraMem INC • Wayne (CA)

On-site
USD 110,000 - 250,000
Machine Learning Engineer
Machine Learning Engineer

Edison Smart® • San Francisco (CA)

On-site
USD 180,000 - 200,000
Medical, dental, and vision coverage
Life, AD&D, and disability insurance
HSA/FSA options
+1
Senior Edge AI Engineer — Embedded ML & Audio
Senior Edge AI Engineer — Embedded ML & Audio

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 110,000 - 250,000
Embedded AI Engineer
Embedded AI Engineer

Bright Vision Technologies • Kirkland (WA)

Remote
USD 100,000 - 150,000
Software Engineer I – Compiler & Runtime
Software Engineer I – Compiler & Runtime

TetraMem Inc • San Jose (CA)

On-site
USD 135,000 - 165,000
Full-time employee benefits
Equity eligibility
Senior Edge AI Engineer — Lead On-Device ML & Edge Systems
Senior Edge AI Engineer — Lead On-Device ML & Edge Systems

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 200,000 - 280,000
Senior Firmware Engineer, Edge AI / NPU Runtime
Senior Firmware Engineer, Edge AI / NPU Runtime

Tacit • San Francisco (CA)

On-site
USD 150,000 - 200,000
Competitive equity package
Comprehensive medical, dental, and vision insurance
Unlimited PTO
+2