Remote AI Accelerator Framework Engineer - Principal

Ampere

Warszawa

On-site

PLN 302,000 - 453,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Premium medical health care
Generous paid time off
Office amenities and snacks
Remote work policy with reimbursement

Job summary

Ampere is a semiconductor design company forging a new era of high-performance, energy-efficient AI compute. As an AI Accelerator Software Principal Engineer – Framework Integration, you will own the strategy and delivery for DL inference across Ampere accelerators, enabling PyTorch/ONNX/llama.cpp integration that drives real-world performance from data centers to edge.

You will lead cross-team efforts, architect end-to-end acceleration across SW/HW, and mentor teams to raise the bar on quality

Qualifications

  • Education & experience: BS CS/CE/EE or related field + 8y; or MS +6y; or PhD +3y.
  • Deep framework expertise: PyTorch, ONNX, and llama.cpp integration with performance focus.
  • Linux accelerator runtime/driver experience (preferred).
  • Strong systems programming + performance tuning: Python and C/C++.

Responsibilities

  • Framework integration leadership for PyTorch/ONNX/llama.cpp into Ampere backend.
  • Drive end-to-end SW/HW acceleration across the stack including inference serving and runtime.
  • Model enablement: optimize performance and accuracy for popular frameworks and stacks.
  • HW/SW co-design to maximize efficiency and scalability across cores and memory.

Skills

Python
C/C++
PyTorch
ONNX
llama.cpp
Performance tuning
Framework integration

Education

BS in Computer Science/Engineering
MS in related field
PhD in related field

Tools

Linux
Linux drivers
Runtime libraries

Job description

Ampere is a semiconductor design company forging a new era of high-performance, energy-efficient AI compute. As an AI Accelerator Software Principal Engineer – Framework Integration, you will own the strategy and delivery for DL inference across Ampere accelerators, enabling PyTorch/ONNX/llama.cpp integration that drives real-world performance from data centers to edge.

You will lead cross-team efforts, architect end-to-end acceleration across SW/HW, and mentor teams to raise the bar on quality

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote AI Accelerator Software Lead — Framework & Performance
Remote AI Accelerator Software Lead — Framework & Performance

Ampere • Warszawa

Hybrid
PLN 416,000 - 625,000
Premium medical health care
Generous paid time off
Office amenities
+2
Senior AI Accelerator Software Architect (Framework & Perf) - Remote
Senior AI Accelerator Software Architect (Framework & Perf) - Remote

Ampere Computing • Warszawa

Hybrid
PLN 417,000 - 625,000
Premium health care
Paid time off
Office amenities
+2
Principal AI Accelerator Runtime Engineer
Principal AI Accelerator Runtime Engineer

Ampere Computing • Warszawa

On-site
PLN 302,000 - 453,000
Premium medical care
Generous PTO
Gym and sauna access
+3
AI Accelerator Software Senior Principal Engineer- Framework Integration
AI Accelerator Software Senior Principal Engineer- Framework Integration

Ampere Computing • Warszawa

Hybrid
PLN 417,000 - 625,000
Premium health care
Paid time off
Office amenities
+2
AI Accelerator Software Senior Principal Engineer- Framework Integration
AI Accelerator Software Senior Principal Engineer- Framework Integration

Ampere • Warszawa

Hybrid
PLN 416,000 - 625,000
Premium medical health care
Generous paid time off
Office amenities
+2
AI Accelerator Software Principal Engineer – Runtime Library
AI Accelerator Software Principal Engineer – Runtime Library

Ampere Computing • Warszawa

On-site
PLN 302,000 - 453,000
Premium medical care
Generous PTO
Gym and sauna access
+3
AI Accelerator Software Principal Engineer- Framework Integration
AI Accelerator Software Principal Engineer- Framework Integration

Ampere • Warszawa

On-site
PLN 302,000 - 453,000
Premium medical health care
Generous paid time off
Office amenities and snacks
+1
AI Accelerator Software Principal Engineer – Runtime Library
AI Accelerator Software Principal Engineer – Runtime Library

Ampere • Warszawa

Hybrid
PLN 301,000 - 453,000
Premium medical health care
Generous paid time off
Office amenities: snacks, gym access
+2
Senior PyTorch Engineer for ML Frameworks
Senior PyTorch Engineer for ML Frameworks

EngineersOfAI • Województwo pomorskie

On-site
PLN 260,000 - 353,000
Senior C++ AI Kernel Engineer (CUDA/GPU)
Senior C++ AI Kernel Engineer (CUDA/GPU)

EPAM Systems • Poland

Hybrid
PLN 240,000 - 360,000
Hybrid by design
Remote work within Poland
Relocation opportunities
+4