Edge AI Engineer: On-Device Models, Distillation & Quantization

Hyphen Connect

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Hyphen Connect seeks an AI Specialist Engineer in San Francisco to enhance AI performance on diverse hardware architectures. You will efficiently develop and deploy cutting-edge AI solutions. Responsibilities include compressing large models for on-device use, developing model distillation pipelines, and benchmarking across NPU/GPU architectures. Ideal candidates should have expertise in AI model optimization, strong C++ and Python skills, and hands-on experience with TensorRT and ONNX Runtime.

Qualifications

  • Expertise in model distillation, pruning, and quantization techniques.
  • Hands-on experience with TensorRT, ONNX Runtime, and edge deployment.
  • Strong abilities in C++ and Python.

Responsibilities

  • Compress and optimize large language and vision models for on-device inference.
  • Develop pipelines for model distillation and hardware-specific compilation.
  • Benchmark performance across NPU/GPU architectures.

Skills

Model distillation
C++
Python
TensorRT
ONNX Runtime

Job description

Hyphen Connect seeks an AI Specialist Engineer in San Francisco to enhance AI performance on diverse hardware architectures. You will efficiently develop and deploy cutting-edge AI solutions. Responsibilities include compressing large models for on-device use, developing model distillation pipelines, and benchmarking across NPU/GPU architectures. Ideal candidates should have expertise in AI model optimization, strong C++ and Python skills, and hands-on experience with TensorRT and ONNX Runtime.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Edge AI Engineer: On-Device Inference & Optimization
Edge AI Engineer: On-Device Inference & Optimization

Hyphen Connect • Seattle (WA)

On-site
USD 100,000 - 140,000
On-Device AI Engineer: Model Optimization & Edge Deployment
On-Device AI Engineer: Model Optimization & Edge Deployment

Hyphen Connect • Boston (MA)

On-site
USD 100,000 - 130,000
AI On-Device Engineer: Model Optimization & Edge Deployment
AI On-Device Engineer: Model Optimization & Edge Deployment

Hyphen Connect • Oregon (WI)

On-site
USD 100,000 - 130,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • Oregon (WI)

On-site
USD 100,000 - 130,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • Seattle (WA)

On-site
USD 100,000 - 140,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • Boston (MA)

On-site
USD 100,000 - 130,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • San Francisco (CA)

On-site
USD 120,000 - 160,000
Gen AI Architect & Strategic Tech Leader
Gen AI Architect & Strategic Tech Leader

Prodapt • San Francisco (CA)

On-site
USD 150,000 - 190,000
Remote AI Inference Engineer — Edge Model Deployment & Optimization
Remote AI Inference Engineer — Edge Model Deployment & Optimization

Quadric Inc. • Burlingame (CA)

On-site
USD 180,000 - 260,000
Health Care Plan (Medical, Dental & Vision)
Retirement Plan (401k, IRA)
Life Insurance (Basic, Voluntary & AD&D)
+7
Staff ML Infrastructure & Performance Engineer
Staff ML Infrastructure & Performance Engineer

Embedding VC • San Mateo (CA)

On-site
USD 120,000 - 150,000