Edge AI Research Engineer: Multimodal Model Compression

Tether

Dublin

On-site

EUR 70,000 - 90,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tether is seeking a member of their AI research team to drive innovation in model compression for advanced multimodal AI systems, focusing on efficient deployment of large language models and vision-language models. The role demands expertise in experimental methods such as quantization, distillation, and pruning.

Candidates should possess a PhD in NLP or Machine Learning and have experience with model compression techniques. This position will enhance the performance of AI systems running on resource-constrained devices in Ireland.

Qualifications

  • PhD in NLP, Machine Learning, or a related field, complemented by a solid track record in AI R&D.
  • Experience with PyTorch deep learning frameworks or equivalent.
  • Hands-on experience with model quantization (QAT and PTQ).
  • Research experience with knowledge distillation and model pruning.

Responsibilities

  • Apply low-bit quantization to reduce model size and latency for generative AI models.
  • Leverage knowledge distillation to enable efficient multimodal reasoning.
  • Implement pruning techniques for reducing computational overhead.
  • Research and apply advanced compression strategies to optimize accuracy-performance balance.
  • Document methodologies and publish findings in top-tier conferences.

Skills

Model Compression
Quantization
Knowledge Distillation
Model Pruning
PyTorch
Neural Network Architectures

Education

PhD in NLP, Machine Learning, or related field
Degree in Computer Science or related field

Job description

Tether is seeking a member of their AI research team to drive innovation in model compression for advanced multimodal AI systems, focusing on efficient deployment of large language models and vision-language models. The role demands expertise in experimental methods such as quantization, distillation, and pruning.

Candidates should possess a PhD in NLP or Machine Learning and have experience with model compression techniques. This position will enhance the performance of AI systems running on resource-constrained devices in Ireland.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Research Engineer (Model Compression & Quantization)
AI Research Engineer (Model Compression & Quantization)

Tether • Dublin

On-site
EUR 70,000 - 90,000
AI Research Engineer: Multi-Modal & LLM Innovation
AI Research Engineer: Multi-Modal & LLM Innovation

Tether • Dublin

On-site
EUR 70,000 - 110,000
AI Research Engineer (Pre-training - LLM & Multi-Modal)
AI Research Engineer (Pre-training - LLM & Multi-Modal)

Tether • Dublin

On-site
EUR 70,000 - 110,000
AI Model Optimization Architect for LLMs & Multimodal
AI Model Optimization Architect for LLMs & Multimodal

Qualcomm • Cork

Hybrid
EUR 120,000 - 180,000
Salary and equity package
Relocation support
Education Assistance
+3
2026 - Senior AI/ML Engineer – Multimodal Content Intelligence - Contractor
2026 - Senior AI/ML Engineer – Multimodal Content Intelligence - Contractor

Huawei Ireland Research Center • Dublin

On-site
EUR 70,000 - 100,000
Cloud AI Performance Engineer
Cloud AI Performance Engineer

Qualcomm • Ireland

On-site
EUR 90,000 - 150,000
Stock bonus
Employee stock purchase scheme
Pension matching scheme
+4
Senior AI Model Optimization Architect for Inference
Senior AI Model Optimization Architect for Inference

Qualcomm • Ireland

On-site
EUR 120,000 - 180,000
Salary, stock and performance related—
Relocation and immigration support
Education Assistance
+1
Cloud AI Inference Performance Engineer
Cloud AI Inference Performance Engineer

Qualcomm • Cork

On-site
EUR 90,000 - 150,000
Stock options
Performance bonus
Relocation assistance
+4
Staff AI Model Optimization Architect — Scalable Inference
Staff AI Model Optimization Architect — Scalable Inference

Qualcomm • Cork

On-site
EUR 150,000 - 190,000
Salary and stock bonus
Relocation assistance
Education assistance
+4
AI Inference Performance Engineer
AI Inference Performance Engineer

Qualcomm • Cork

Hybrid
EUR 90,000 - 130,000
Salary review and performance bonus
Relocation support
Education Assistance
+2