Edge AI Engineer: On-Device Inference & Optimization
Hyphen Connect
Seattle (WA)
On-site
USD 100,000 - 140,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Job summary
Hyphen Connect, based in Seattle, is seeking an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. The role involves developing and deploying cutting-edge AI solutions across diverse hardware architectures. Ideal candidates will have expertise in model distillation, pruning, and quantization techniques, along with strong skills in C++ and Python. This position promises significant challenges and rewards in the evolving field of AI.
Qualifications
Expertise in model distillation, pruning, and quantization techniques.
Hands-on experience with TensorRT and ONNX Runtime.
Strong skills in C++ and Python.
Responsibilities
Compress and optimize models for on-device inference.
Develop pipelines for model distillation and compilation.
Benchmark performance across various architectures.
Skills
Model distillation
C++
Python
4-bit/8-bit quantization
TensorRT
ONNX Runtime
Hardware-specific compilation
Job description
Hyphen Connect, based in Seattle, is seeking an AI Specialist Engineer to enhance the performance of large language and vision models for on-device inference. The role involves developing and deploying cutting-edge AI solutions across diverse hardware architectures. Ideal candidates will have expertise in model distillation, pruning, and quantization techniques, along with strong skills in C++ and Python. This position promises significant challenges and rewards in the evolving field of AI.