Remote AI Research Engineer: Model Compression & Multimodal

Tether.io

Ireland

On-site

EUR 90,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Tether.io is seeking an AI research engineer to advance compression techniques for multimodal models, including LLMs and VLMs. You will focus on reducing model footprint and deployment costs while preserving accuracy, enabling efficient edge deployment and scalable AI services.

The role emphasizes quantization, distillation, and pruning, with opportunities to publish findings and contribute to cutting-edge fintech AI research.

Qualifications

  • PhD in NLP/ML or related field with strong publications.
  • Experience with PyTorch and deep learning.
  • Hands-on with model compression (quantization, distillation, pruning).

Responsibilities

  • Apply low-bit quantization to reduce model size and latency for multimodal models while maintaining accuracy.
  • Leverage knowledge distillation to transfer capabilities from larger models to smaller ones.
  • Implement pruning techniques to remove redundant parameters and attention heads.
  • Analyze trade-offs between model efficiency and accuracy; propose improvements.
  • Document methodologies and results to support reproducibility.

Skills

English communication
Multimodal AI research
Edge deployment
Quantization-aware training
Post-training quantization

Education

PhD in NLP/ML or related field
Bachelor's in Computer Science or related field

Tools

PyTorch
C++
Knowledge distillation
Model pruning
Quantization
Transformers

Job description

Tether.io is seeking an AI research engineer to advance compression techniques for multimodal models, including LLMs and VLMs. You will focus on reducing model footprint and deployment costs while preserving accuracy, enabling efficient edge deployment and scalable AI services.

The role emphasizes quantization, distillation, and pruning, with opportunities to publish findings and contribute to cutting-edge fintech AI research.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Edge AI Research Engineer: Multimodal Model Compression
Edge AI Research Engineer: Multimodal Model Compression

Tether • Dublin

On-site
EUR 70,000 - 90,000
AI Research Engineer (Model Compression & Quantization)
AI Research Engineer (Model Compression & Quantization)

Tether • Dublin

On-site
EUR 70,000 - 90,000
AI Research Engineer: Multi-Modal & LLM Innovation
AI Research Engineer: Multi-Modal & LLM Innovation

Tether • Dublin

On-site
EUR 70,000 - 110,000
AI Research Engineer (Pre-training - LLM & Multi-Modal)
AI Research Engineer (Pre-training - LLM & Multi-Modal)

Tether • Dublin

On-site
EUR 70,000 - 110,000
AI Model Optimization Architect for LLMs & Multimodal
AI Model Optimization Architect for LLMs & Multimodal

Qualcomm • Cork

Hybrid
EUR 120,000 - 180,000
Salary and equity package
Relocation support
Education Assistance
+3
Senior AI Model Optimization Architect for Inference
Senior AI Model Optimization Architect for Inference

Qualcomm • Ireland

On-site
EUR 120,000 - 180,000
Salary, stock and performance related—
Relocation and immigration support
Education Assistance
+1
Staff AI Model Optimization Architect — Scalable Inference
Staff AI Model Optimization Architect — Scalable Inference

Qualcomm • Cork

On-site
EUR 150,000 - 190,000
Salary and stock bonus
Relocation assistance
Education assistance
+4
AI Inference Engineer — High-Performance, Low-Latency ML
AI Inference Engineer — High-Performance, Low-Latency ML

F5 • Dublin

On-site
EUR 90,000 - 150,000
Senior AI Inference Engineer - High-Throughput LLM Serving
Senior AI Inference Engineer - High-Throughput LLM Serving

Confidential • Ireland

On-site
EUR 120,000 - 180,000
2026 - Senior AI/ML Engineer – Multimodal Content Intelligence - Contractor
2026 - Senior AI/ML Engineer – Multimodal Content Intelligence - Contractor

Huawei Ireland Research Center • Dublin

On-site
EUR 70,000 - 100,000