Edge AI Engineer: On-Device Inference & Optimization

Hyphen Connect

Seattle (WA)

On-site

USD 100,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Hyphen Connect, based in Seattle, is seeking an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. The role involves developing and deploying cutting-edge AI solutions across diverse hardware architectures. Ideal candidates will have expertise in model distillation, pruning, and quantization techniques, along with strong skills in C++ and Python. This position promises significant challenges and rewards in the evolving field of AI.

Qualifications

  • Expertise in model distillation, pruning, and quantization techniques.
  • Hands-on experience with TensorRT and ONNX Runtime.
  • Strong skills in C++ and Python.

Responsibilities

  • Compress and optimize models for on-device inference.
  • Develop pipelines for model distillation and compilation.
  • Benchmark performance across various architectures.

Skills

Model distillation
C++
Python
4-bit/8-bit quantization
TensorRT
ONNX Runtime
Hardware-specific compilation

Job description

Hyphen Connect, based in Seattle, is seeking an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. The role involves developing and deploying cutting-edge AI solutions across diverse hardware architectures. Ideal candidates will have expertise in model distillation, pruning, and quantization techniques, along with strong skills in C++ and Python. This position promises significant challenges and rewards in the evolving field of AI.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI On-Device Engineer: Model Optimization & Edge Deployment
AI On-Device Engineer: Model Optimization & Edge Deployment

Hyphen Connect • Oregon (WI)

On-site
USD 100,000 - 130,000
Edge AI Engineer: On-Device Models, Distillation & Quantization
Edge AI Engineer: On-Device Models, Distillation & Quantization

Hyphen Connect • San Francisco (CA)

On-site
USD 120,000 - 160,000
On-Device AI Engineer: Model Optimization & Edge Deployment
On-Device AI Engineer: Model Optimization & Edge Deployment

Hyphen Connect • Boston (MA)

On-site
USD 100,000 - 130,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • Oregon (WI)

On-site
USD 100,000 - 130,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • Boston (MA)

On-site
USD 100,000 - 130,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • Seattle (WA)

On-site
USD 100,000 - 140,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • San Francisco (CA)

On-site
USD 120,000 - 160,000
Remote AI Inference Engineer — Edge Model Deployment & Optimization
Remote AI Inference Engineer — Edge Model Deployment & Optimization

Quadric Inc. • Burlingame (CA)

On-site
USD 180,000 - 260,000
Health Care Plan (Medical, Dental & Vision)
Retirement Plan (401k, IRA)
Life Insurance (Basic, Voluntary & AD&D)
+7
AI Inference Architect & Performance Engineer
AI Inference Architect & Performance Engineer

ElastixAI INC. • Seattle (WA)

Hybrid
Competitive compensation
Comprehensive medical, dental, and vision coverage
Flexible Time Off (FTO)
+4
Senior AI Inference Engineer — Edge ML, Porting Models
Senior AI Inference Engineer — Edge ML, Porting Models

quadric.io, Inc • Burlingame (CA)

On-site
USD 120,000 - 150,000
Health Care Plan (Medical, Dental & Vision)
Retirement Plan (401k, IRA)
Life Insurance (Basic, Voluntary & AD&D)
+7