Edge AI Model Compression & Quantization Engineer

Tether.io

United Kingdom

On-site

GBP 60,000 - 90,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tether.io is seeking an AI Research Scientist to innovate in model compression and efficient deployment for advanced multimodal AI systems. The successful candidate will focus on reducing model footprint while maintaining performance across edge devices.

Your responsibilities will include employing quantization, knowledge distillation, and pruning techniques to streamline AI models, and you'll have the opportunity to publish your research in top-tier conferences.

Qualifications

  • PhD in NLP, Machine Learning, or related field with publications.
  • Experience with PyTorch and model quantization.
  • Hands-on experience with knowledge distillation and model pruning.

Responsibilities

  • Apply low-bit quantization to reduce model size and latency.
  • Leverage knowledge distillation for efficient multimodal reasoning.
  • Implement pruning techniques to reduce computational overhead.

Skills

Model compression techniques
Quantization
Knowledge distillation
Model pruning
Experience with PyTorch
Neural network architectures
Familiarity with C++

Education

PhD in NLP or Machine Learning
Degree in Computer Science or related field

Job description

Tether.io is seeking an AI Research Scientist to innovate in model compression and efficient deployment for advanced multimodal AI systems. The successful candidate will focus on reducing model footprint while maintaining performance across edge devices.

Your responsibilities will include employing quantization, knowledge distillation, and pruning techniques to streamline AI models, and you'll have the opportunity to publish your research in top-tier conferences.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Edge AI Scientist: Compress & Deploy Multimodal Models
Edge AI Scientist: Compress & Deploy Multimodal Models

Eworker • United Kingdom

On-site
GBP 90,000 - 130,000
Edge AI Inference Engineer: Kernel & Serving
Edge AI Inference Engineer: Kernel & Serving

Tether • United Kingdom

Remote
GBP 80,000 - 100,000
AI Research Engineer — LLM & Multimodal Architectures
AI Research Engineer — LLM & Multimodal Architectures

Tether.io • United Kingdom

On-site
GBP 80,000 - 120,000
Remote AI Research Engineer — LLM & Multimodal
Remote AI Research Engineer — LLM & Multimodal

Tether • United Kingdom

Remote
GBP 111,000 - 141,000
Remote AI Research Engineer - LLM & Multimodal Architect
Remote AI Research Engineer - LLM & Multimodal Architect

Tether.io • United Kingdom

On-site
GBP 90,000 - 150,000
AI Research Engineer, Multimodal & LLM Architect
AI Research Engineer, Multimodal & LLM Architect

Tether Operations Limited • United Kingdom

On-site
GBP 70,000 - 110,000
AI Research Engineer (Pre-training - LLM & Multi-Modal) - 100% Remote Worldwide
AI Research Engineer (Pre-training - LLM & Multi-Modal) - 100% Remote Worldwide

Tether.io • United Kingdom

On-site
GBP 90,000 - 150,000
AI Research Engineer (Pre-training - LLM & Multi-Modal) - 100% Remote Worldwide
AI Research Engineer (Pre-training - LLM & Multi-Modal) - 100% Remote Worldwide

Tether • United Kingdom

Remote
GBP 111,000 - 141,000
Remote AI Vision-Language Research Engineer
Remote AI Vision-Language Research Engineer

Tether • Greater London

On-site
GBP 60,000 - 80,000
Work remotely with global talent
Opportunity for innovation in fintech
Collaboration with experts in the field
AI Computer Vision Engineer — Edge-to-Cloud SDKs
AI Computer Vision Engineer — Edge-to-Cloud SDKs

Zebra Technologies Europe Limited • Greater London

On-site
GBP 75,000 - 110,000
Annual incentive 12% of base pay