Triton AI Compiler Architect for NPU Performance

Adecco

Warszawa

On-site

PLN 280,000 - 420,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

On-site in Warsaw

Job summary

Adecco is seeking a highly capable engineer to join our deep learning infrastructure team in Warsaw, Poland. You will own the Triton compiler and kernel optimization framework for NPUs, driving AI workloads with advanced performance engineering and compiler techniques.

You will implement state-of-the-art Triton kernels, optimize memory patterns, and ensure tight integration with PyTorch and other runtimes. This role combines AI compilation, NPU programming, and system performance, with mentoring

Qualifications

  • Master’s or PhD in a relevant field and 5+ years in NPU/GPU programming or compiler development.
  • Strong knowledge of accelerator architectures and performance bottlenecks.
  • Proficiency in Triton, PTX, or LLVM IR for low-level optimization.

Responsibilities

  • Lead design and development of the Triton compiler and performance framework for NPUs.
  • Implement high-performance Triton kernels (Attention, MatMul, LayerNorm, Conv, Softmax).
  • Optimize memory access, scheduling, and cache usage for scalable performance.
  • Integrate Triton with PyTorch, XLA, and runtime stacks for end-to-end performance.
  • Apply auto-tuning, kernel fusion, and operator scheduling to maximize throughput.
  • Mentor teammates and establish processes for performance analysis and optimization.
  • Stay current with MLIR, TVM, and related compiler tech to push boundaries.
  • Perform workload fingerprinting for large models to guide optimizations.

Skills

NPU programming
Kernel optimization
Performance profiling
System design
Mentoring

Education

Master’s degree in Computer Architecture or related field

Tools

Triton
LLVM IR
PTX
TVM
MLIR
Cutlass

Job description

Adecco is seeking a highly capable engineer to join our deep learning infrastructure team in Warsaw, Poland. You will own the Triton compiler and kernel optimization framework for NPUs, driving AI workloads with advanced performance engineering and compiler techniques.

You will implement state-of-the-art Triton kernels, optimize memory patterns, and ensure tight integration with PyTorch and other runtimes. This role combines AI compilation, NPU programming, and system performance, with mentoring

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Lead AI Compiler & Kernel Optimization for NPUs
Lead AI Compiler & Kernel Optimization for NPUs

microTECH Global LTD • Warszawa

On-site
PLN 260,000 - 380,000
Compiler Expert
Compiler Expert

Adecco • Warszawa

On-site
PLN 280,000 - 420,000
On-site in Warsaw
Compiler Engineer
Compiler Engineer

microTECH Global LTD • Warszawa

On-site
PLN 260,000 - 380,000
ML Frameworks Engineer (Triton) – AI Hardware & Software
ML Frameworks Engineer (Triton) – AI Hardware & Software

Graphcore • Województwo pomorskie

On-site
PLN 180,000 - 240,000
Flexible working
Generous leave
Pension matching
+4
Software Engineer - Triton Gdańsk, Pomeranian Voivodeship, Poland
Software Engineer - Triton Gdańsk, Pomeranian Voivodeship, Poland

Graphcore • Województwo pomorskie

Hybrid
PLN 120,000 - 180,000
Flexible working
Healthcare & dental cover
Phantom equity
+4
ML Frameworks Engineer for AI Hardware (Triton/PyTorch)
ML Frameworks Engineer for AI Hardware (Triton/PyTorch)

Graphcore • Poland

Hybrid
PLN 150,000 - 210,000
Flexible working
Healthcare and dental
Phantom equity
+4
Senior Compiler Performance Engineer: Optimize AI/CPU
Senior Compiler Performance Engineer: Optimize AI/CPU

AMD • Warszawa

On-site
PLN 100,000 - 130,000
Senior Deep Learning Compiler Engineer - PyTorch
Senior Deep Learning Compiler Engineer - PyTorch

NVIDIA • Warszawa

On-site
PLN 293,000 - 507,000
Staff AI Compute Kernels Lead - High-Performance
Staff AI Compute Kernels Lead - High-Performance

Cerebras • Województwo pomorskie

On-site
PLN 350,000 - 475,000
Annual leave policy
Medical and dental health plans
Gym card
+1
Staff AI Compute Libraries & Performance Engineer
Staff AI Compute Libraries & Performance Engineer

Graphcore • Województwo pomorskie

On-site
PLN 351,000 - 474,000
Flexible working
Comprehensive healthcare and dental
Phantom equity
+3