Neuronetwork System Engineer

TetraMem - Accelerate The World

San Jose (CA)

On-site

USD 110,000 - 250,000

Full time

3 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

TetraMem - Accelerate The World is seeking a senior ML engineer to design and optimize edge AI models for audio processing on embedded platforms, including FPGA and ASIC. You will collaborate with hardware and software teams to deploy production-ready ML solutions.

The role emphasizes research-driven model efficiency, latency reduction, and power optimization, with opportunities to mentor junior engineers and publish findings.

Qualifications

  • 5+ years of experience or PhD in Computer Science, Electrical Engineering, or related fields.
  • Strong experience in machine learning, with a focus on edge AI and lightweight model deployment.
  • Expertise in ML frameworks such as PyTorch, TensorFlow, JAX.
  • Proficiency in programming languages such as C/C++, Python, and experience with ML model optimization.

Responsibilities

  • Develop, optimize, and deploy lightweight machine learning models for edge AI applications, particularly for audio processing.
  • Implement and optimize ML models on embedded platforms, including FPGA and custom ASIC solutions.
  • Work closely with hardware and software teams to integrate ML models into production systems.
  • Research and implement state-of-the-art ML techniques to enhance model efficiency, latency, and power consumption for embedded AI applications.
  • Improve inference efficiency and model compression techniques, including quantization, pruning, and knowledge distillation.
  • Collaborate with cross-functional teams to drive innovation and contribute to the overall system architecture.
  • Provide technical leadership and mentorship to junior engineers.
  • Publish research findings, present at conferences, and contribute to open-source projects when applicable.

Skills

Edge AI
PyTorch
TensorFlow
JAX
C/C++
Python
Model optimization

Education

PhD or MS in Computer Science or Electrical Engineering

Tools

TensorRT
ONNX
TFLite/LiteRT
ncnn
CoreML

Job description

  • Develop, optimize, and deploy lightweight machine learning models for edge AI applications, particularly for audio processing.
  • Implement and optimize ML models on embedded platforms, including FPGA and custom ASIC solutions.
  • Work closely with hardware and software teams to integrate ML models into production systems.
  • Research and implement state-of-the-art ML techniques to enhance model efficiency, latency, and power consumption for embedded AI applications.
  • Improve inference efficiency and model compression techniques, including quantization, pruning, and knowledge distillation.
  • Collaborate with cross-functional teams to drive innovation and contribute to the overall system architecture.
  • Provide technical leadership and mentorship to junior engineers.
  • Publish research findings, present at conferences, and contribute to open-source projects when applicable.
Responsibilities
  • Develop, optimize, and deploy lightweight machine learning models for edge AI applications, particularly for audio processing.
  • Implement and optimize ML models on embedded platforms, including FPGA and custom ASIC solutions.
  • Work closely with hardware and software teams to integrate ML models into production systems.
  • Research and implement state-of-the-art ML techniques to enhance model efficiency, latency, and power consumption for embedded AI applications.
  • Improve inference efficiency and model compression techniques, including quantization, pruning, and knowledge distillation.
  • Collaborate with cross-functional teams to drive innovation and contribute to the overall system architecture.
  • Provide technical leadership and mentorship to junior engineers.
  • Publish research findings, present at conferences, and contribute to open-source projects when applicable.
Requirements
  • 5+ years of experience or PhD in Computer Science, Electrical Engineering, or related fields.
  • Strong experience in machine learning, with a focus on edge AI and lightweight model deployment.
  • Expertise in ML frameworks such as PyTorch, TensorFlow, JAX.
  • Proficiency in programming languages such as C/C++, Python, and experience with ML model optimization.
  • Ability to work independently and collaboratively in a fast-paced startup environment.
  • Ability to provide mentorship, technical guidance, and career development support to junior engineers and interns.
Experience in one or more of the following areas considered a strong plus:
  • Understanding of ML compiler and runtime design.
  • Experience working with tools such as Optimum, ONNX, TensorRT, TFLite/LiteRT, ncnn, or CoreML.
  • Familiarity with hardware acceleration techniques.
  • Experience in embedded system development.

Salary Range: $110,000 - $250,000 / year

TetraMem celebrates diversity and is committed to creating an inclusive environment for all employees. We are proud to be an Equal Opportunity Employer and welcome applicants from all backgrounds. Qualified candidates will receive consideration for employment without regard to race, color, religion, creed, sex, gender identity or expression, sexual orientation, national origin, ancestry, age, marital status, medical condition, disability, genetic information, military or veteran status, or any other characteristic protected by applicable federal, state, or local law.

TetraMem is committed to providing reasonable accommodations to qualified applicants with disabilities throughout the recruitment process. Applicants requiring accommodation may contact Human Resources for assistance.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Neuronetwork System Engineer
Neuronetwork System Engineer

TetraMem Inc. • San Jose (CA)

Hybrid
USD 110,000 - 250,000
Neuronetwork System Engineer
Neuronetwork System Engineer

TetraMem INC • Wayne (CA)

On-site
USD 110,000 - 250,000
Neuronetwork System Engineer
Neuronetwork System Engineer

TETRAMEM INC • San Jose (CA)

On-site
USD 110,000 - 250,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 200,000 - 280,000
Field Application Engineer
Field Application Engineer

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 120,000 - 180,000
Software Engineer I – Compiler & Runtime
Software Engineer I – Compiler & Runtime

TetraMem Inc • San Jose (CA)

On-site
USD 135,000 - 165,000
Full-time employee benefits
Equity eligibility
US 2026 Software - Machine Learning Intern
US 2026 Software - Machine Learning Intern

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 48,000 - 62,000
Software Engineer I – Compiler & Runtime
Software Engineer I – Compiler & Runtime

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 135,000 - 165,000
Lead Edge ML Engineer – Audio & Embedded Inference
Lead Edge ML Engineer – Audio & Embedded Inference

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 200,000 - 280,000
Edge AI System Engineer — Lightweight ML on Embedded
Edge AI System Engineer — Lightweight ML on Embedded

TetraMem - Accelerate The World • San Jose (CA)

On-site
USD 110,000 - 250,000