Senior AI Model Optimization Architect for Inference

Qualcomm

Ireland

On-site

EUR 120,000 - 180,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Salary, stock and performance related—
Relocation and immigration support
Education Assistance
Life, Medical, Income and Travel保险

Job summary

QT Technologies Ireland Limited is seeking a Staff Engineer – AI Model Optimization Architect to lead end-to-end transformation and optimization for LLMs, VLMs, diffusion and multimodal models on Qualcomm accelerators. You will collaborate with compiler, performance, and accuracy teams to translate models into accelerator-efficient execution, balancing throughput, latency, memory and quality.

The role covers Day0 enablement through production deployment, with emphasis on scaling optimizations

Qualifications

  • Expert level expertise in PyTorch and inference-focused model optimization.
  • Hands-on experience with graph capture and compilation workflows (TorchDynamo, torch.compile).
  • Deep understanding of transformer architectures, attention, MVEs, and performance trade-offs.

Responsibilities

  • Architect and deliver model optimization strategies transforming PyTorch models for Qualcomm accelerators.
  • Drive graph capture and deployment using PyTorch, ONNX, and torch.compile with model rewrites.
  • Design and implement fusion kernels using DSL-based approaches (e.g., Triton).
  • Co-design lowering strategies, kernel fusion, layout decisions, and runtime integration with compiler teams.
  • Profile and optimize LLM/VLM/diffusion inference for throughput and latency across modes.

Skills

PyTorch
Python
TorchDynamo
ONNX
Triton
Model optimization
Inference engineering
Distributed systems

Education

Bachelor's degree in Engineering/CS/related (Bachelors)
Master's degree in Engineering/CS/related
PhD in Engineering/CS/related

Tools

TorchDynamo
ONNX
Triton

Job description

QT Technologies Ireland Limited is seeking a Staff Engineer – AI Model Optimization Architect to lead end-to-end transformation and optimization for LLMs, VLMs, diffusion and multimodal models on Qualcomm accelerators. You will collaborate with compiler, performance, and accuracy teams to translate models into accelerator-efficient execution, balancing throughput, latency, memory and quality.

The role covers Day0 enablement through production deployment, with emphasis on scaling optimizations

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Model Optimization Architect for LLMs & Multimodal
AI Model Optimization Architect for LLMs & Multimodal

Qualcomm • Cork

Hybrid
EUR 120,000 - 180,000
Salary and equity package
Relocation support
Education Assistance
+3
AI Model Optimization Architect - Cork, Ireland
AI Model Optimization Architect - Cork, Ireland

Qualcomm • Ireland

On-site
EUR 120,000 - 180,000
Salary, stock and performance related—
Relocation and immigration support
Education Assistance
+1
AI Model Optimization Architect - Cork, Ireland Cork, Ireland Software Engineering Posted a day ago
AI Model Optimization Architect - Cork, Ireland Cork, Ireland Software Engineering Posted a day ago

Qualcomm • Cork

Hybrid
EUR 120,000 - 180,000
Salary and equity package
Relocation support
Education Assistance
+3
Senior AI Systems Software Engineer
Senior AI Systems Software Engineer

Qualcomm • Ireland

On-site
EUR 120,000 - 180,000
Stock bonus
Maternity/Paternity Leave
Employee stock purchase scheme
+1
AI Performance Engineer (Cloud AI Engineering), Sr | Staff | Sr. Staff - Cork, Ireland
AI Performance Engineer (Cloud AI Engineering), Sr | Staff | Sr. Staff - Cork, Ireland

Qualcomm • Ireland

On-site
EUR 90,000 - 150,000
Stock bonus
Employee stock purchase scheme
Pension matching scheme
+4
Cloud AI Performance Engineer
Cloud AI Performance Engineer

Qualcomm • Ireland

On-site
EUR 90,000 - 150,000
Stock bonus
Employee stock purchase scheme
Pension matching scheme
+4
Senior AI Inference Engineer - High-Throughput LLM Serving
Senior AI Inference Engineer - High-Throughput LLM Serving

Confidential • Ireland

On-site
EUR 120,000 - 180,000
Principal Cloud AI Engineer — LLM Serving
Principal Cloud AI Engineer — LLM Serving

Qualcomm • Cork

Hybrid
EUR 110,000 - 160,000
Salary and stock bonus
Relocation and immigration support
Education Assistance
+1
LLM Serving Engineer (Cloud AI Engineering), Senior / Staff Engineer - Cork, Ireland Cork, Ireland Software Engineering
LLM Serving Engineer (Cloud AI Engineering), Senior / Staff Engineer - Cork, Ireland Cork, Ireland Software Engineering

Qualcomm • Cork

Hybrid
EUR 110,000 - 160,000
Salary and stock options
Performance bonus
Maternity/Paternity Leave
+8
LLM Serving Engineer (Cloud AI Engineering), Senior / Staff Engineer - Cork, Ireland
LLM Serving Engineer (Cloud AI Engineering), Senior / Staff Engineer - Cork, Ireland

Qualcomm • Ireland

On-site
EUR 110,000 - 170,000
Salary and bonus
Parental leave
Employee stock purchase
+4