AI On-Device Engineer: Model Optimization & Edge Deployment

Hyphen Connect

Oregon (WI)

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Hyphen Connect is seeking an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. You will be crucial in developing AI solutions, optimizing model performance across diverse hardware architectures. Key qualifications include expertise in model distillation and quantization techniques, hands-on experience with TensorRT and ONNX Runtime, along with strong C++ and Python skills. This role offers exciting opportunities to work on cutting-edge technology in the United States.

Qualifications

  • Expertise in model distillation, pruning, and quantization techniques.
  • Hands-on experience with TensorRT, ONNX Runtime, and edge deployment.
  • Strong C++ and Python skills.

Responsibilities

  • Compress and optimize large language and vision models for on-device inference.
  • Develop pipelines for model distillation and hardware-specific compilation.
  • Benchmark performance across various NPU/GPU architectures.

Skills

Model distillation
Pruning techniques
Quantization techniques
TensorRT
ONNX Runtime
C++
Python

Job description

Hyphen Connect is seeking an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. You will be crucial in developing AI solutions, optimizing model performance across diverse hardware architectures. Key qualifications include expertise in model distillation and quantization techniques, hands-on experience with TensorRT and ONNX Runtime, along with strong C++ and Python skills. This role offers exciting opportunities to work on cutting-edge technology in the United States.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

On-Device AI Engineer: Model Optimization & Edge Deployment
On-Device AI Engineer: Model Optimization & Edge Deployment

Hyphen Connect • Boston (MA)

On-site
USD 100,000 - 130,000
Edge AI Engineer: On-Device Inference & Optimization
Edge AI Engineer: On-Device Inference & Optimization

Hyphen Connect • Seattle (WA)

On-site
USD 100,000 - 140,000
Edge AI Engineer: On-Device Models, Distillation & Quantization
Edge AI Engineer: On-Device Models, Distillation & Quantization

Hyphen Connect • San Francisco (CA)

On-site
USD 120,000 - 160,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • Boston (MA)

On-site
USD 100,000 - 130,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • Oregon (WI)

On-site
USD 100,000 - 130,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • Seattle (WA)

On-site
USD 100,000 - 140,000
AI Specialist (AI Engineering)
AI Specialist (AI Engineering)

Hyphen Connect • San Francisco (CA)

On-site
USD 120,000 - 160,000
Remote AI Inference Engineer — Edge Model Deployment & Optimization
Remote AI Inference Engineer — Edge Model Deployment & Optimization

Quadric Inc. • Burlingame (CA)

On-site
USD 180,000 - 260,000
Health Care Plan (Medical, Dental & Vision)
Retirement Plan (401k, IRA)
Life Insurance (Basic, Voluntary & AD&D)
+7
Senior AI Inference Engineer — Edge ML, Porting Models
Senior AI Inference Engineer — Edge ML, Porting Models

quadric.io, Inc • Burlingame (CA)

On-site
USD 120,000 - 150,000
Health Care Plan (Medical, Dental & Vision)
Retirement Plan (401k, IRA)
Life Insurance (Basic, Voluntary & AD&D)
+7
Remote Edge AI Engineer - On-Device ML & Optimization
Remote Edge AI Engineer - On-Device ML & Optimization

Visa Hunt • United States

On-site
USD 100,000 - 155,000