AI Model Optimization Architect for LLMs & Multimodal

Qualcomm

Cork

Hybrid

EUR 120,000 - 180,000

Full time

2 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Salary and equity package
Relocation support
Education Assistance
Pension scheme
Health and life insurance
Employee clubs

Job summary

QT Technologies Ireland Limited in Cork is seeking a Staff Engineer – AI Model Optimization Architect to lead end-to-end model transformation for LLMs, VLMs, diffusion, and multimodal models on Qualcomm accelerators.

You will collaborate with compiler, performance, and accuracy teams to optimize graphs, implement fusion kernels, and scale distributed inference across multi-core systems.

Qualifications

  • Expert level expertise in PyTorch inference optimization.
  • Hands on experience with torch.compile / TorchDynamo workflows.
  • Deep understanding of transformer architectures and attention mechanisms.
  • Practical experience with KVcache behavior and serving time optimizations.
  • Strong foundation in computer architecture, ML accelerators and distributed systems.
  • Proven ability to lead cross-functional technical efforts and influence design decisions.
  • Academic background with relevant advanced degree or equivalent experience.

Responsibilities

  • Architect and deliver model optimization strategies transforming PyTorch models for efficient inference on Qualcomm accelerators.
  • Drive graph capture and deployment using PyTorch, ONNX, and torch.compile with model rewrites.
  • Design and implement fusion kernels using DSL-based approaches (e.g., Triton) for performance critical ops.
  • Co-design lowering strategies, kernel fusion, layouts, and runtime integration with teams.
  • Profile and optimize LLM/VLM/diffusion inference for throughput and latency across batched inputs.
  • Manage KVcache behavior and decoding strategies for long-context performance.
  • Enable and optimize continuous batching and distributed inference across multi-core/multi-device systems.
  • Scale model optimizations to new hardware architectures with reusable patterns and tooling.

Skills

PyTorch
Python
TorchDynamo
ONNX
Transformer architectures
KVcache
Distributed systems
C++/Python integration

Education

Bachelor's degree in Engineering, Information Systems, Computer Science
Master's degree in Engineering, Information Systems, Computer Science
PhD in Engineering, Information Systems, Computer Science

Tools

Triton
TorchDynamo tooling

Job description

QT Technologies Ireland Limited in Cork is seeking a Staff Engineer – AI Model Optimization Architect to lead end-to-end model transformation for LLMs, VLMs, diffusion, and multimodal models on Qualcomm accelerators.

You will collaborate with compiler, performance, and accuracy teams to optimize graphs, implement fusion kernels, and scale distributed inference across multi-core systems.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Model Optimization Architect for Inference
Senior AI Model Optimization Architect for Inference

Qualcomm • Ireland

On-site
EUR 120,000 - 180,000
Salary, stock and performance related—
Relocation and immigration support
Education Assistance
+1
AI Model Optimization Architect - Cork, Ireland
AI Model Optimization Architect - Cork, Ireland

Qualcomm • Ireland

On-site
EUR 120,000 - 180,000
Salary, stock and performance related—
Relocation and immigration support
Education Assistance
+1
AI Model Optimization Architect - Cork, Ireland Cork, Ireland Software Engineering Posted a day ago
AI Model Optimization Architect - Cork, Ireland Cork, Ireland Software Engineering Posted a day ago

Qualcomm • Cork

Hybrid
EUR 120,000 - 180,000
Salary and equity package
Relocation support
Education Assistance
+3
Principal Cloud AI Engineer — LLM Serving
Principal Cloud AI Engineer — LLM Serving

Qualcomm • Cork

Hybrid
EUR 110,000 - 160,000
Salary and stock bonus
Relocation and immigration support
Education Assistance
+1
Senior AI Systems Software Engineer
Senior AI Systems Software Engineer

Qualcomm • Ireland

On-site
EUR 120,000 - 180,000
Stock bonus
Maternity/Paternity Leave
Employee stock purchase scheme
+1
AI Performance Engineer (Cloud AI Engineering), Sr | Staff | Sr. Staff - Cork, Ireland
AI Performance Engineer (Cloud AI Engineering), Sr | Staff | Sr. Staff - Cork, Ireland

Qualcomm • Ireland

On-site
EUR 90,000 - 150,000
Stock bonus
Employee stock purchase scheme
Pension matching scheme
+4
LLM Serving Engineer (Cloud AI Engineering), Senior / Staff Engineer - Cork, Ireland Cork, Ireland Software Engineering
LLM Serving Engineer (Cloud AI Engineering), Senior / Staff Engineer - Cork, Ireland Cork, Ireland Software Engineering

Qualcomm • Cork

Hybrid
EUR 110,000 - 160,000
Salary and stock options
Performance bonus
Maternity/Paternity Leave
+8
LLM Serving Engineer (Cloud AI Engineering), Senior / Staff Engineer - Cork, Ireland
LLM Serving Engineer (Cloud AI Engineering), Senior / Staff Engineer - Cork, Ireland

Qualcomm • Ireland

On-site
EUR 110,000 - 170,000
Salary and bonus
Parental leave
Employee stock purchase
+4
Senior AI Inference Engineer - High-Throughput LLM Serving
Senior AI Inference Engineer - High-Throughput LLM Serving

Confidential • Ireland

On-site
EUR 120,000 - 180,000
Cloud AI Performance Engineer
Cloud AI Performance Engineer

Qualcomm • Ireland

On-site
EUR 90,000 - 150,000
Stock bonus
Employee stock purchase scheme
Pension matching scheme
+4