AI Accelerator Software Principal Engineer – Runtime Library

Ampere Computing

Warszawa

On-site

PLN 302,000 - 453,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Premium medical care
Generous PTO
Gym and sauna access
Flexible hours
Remote work reimbursement
Sports card

Job summary

Ampere Computing is hiring an AI Accelerator Principal Software Engineer to lead the design and optimization of AI runtime software for its deep learning accelerators. You will shape execution, memory management, and orchestration across the SW/HW stack to deliver high throughput and low latency for diverse model types and frameworks.

You will collaborate with hardware and platform teams to ensure robust deployment, performance verification, and scalable integration in production workloads.

Qualifications

  • Strong background in developing user‑mode drivers and runtime libraries for GPUs or DL accelerators on Linux or RTOS.
  • Hands‑on experience with AI frameworks and runtime integration (e.g., PyTorch, ONNX).
  • Excellent C/C++ systems programming, memory management, threading, and performance profiling.

Responsibilities

  • Design, develop, and optimize an AI Runtime Library for Ampere accelerators to support diverse models and frameworks.
  • Own end‑to‑end acceleration paths across the SW/HW stack, including inference serving, graph/IR execution, and memory management.
  • Drive HW/SW co‑design and optimization for throughput, latency, and memory efficiency.
  • Collaborate with hardware and systems teams to ensure runtime alignment with accelerator capabilities.
  • Integrate runtime components into Ampere platform stacks for production‑like workloads.

Skills

C/C++
Systems programming
Performance profiling
AI framework enablement
PyTorch
llama.cpp
ONNX
GPU runtime libraries

Education

BS in CS/CE/EE or software eng
MS in CS/CE/EE or software eng
PhD in relevant field

Tools

Linux
RTOS

Job description

AI Accelerator Software Principal Engineer – Runtime Library

Ampere is a semiconductor design company for a new era, leading the future of computing with an innovative approach to CPU design focused on high-performance, energy efficient AI compute.

As a pioneer in the new frontier of energy efficient high-performance computing, Ampere is part of the Softbank Group of companies driving sustainable computing for AI, Cloud, and edge applications.

Join us at Ampere and work alongside a passionate and growing team.

About the Role

As an AI Accelerator Principal Software Engineer – Runtime Library, you will lead the design, development, and optimization of AI runtime software that enables multiple state-of-the-art deep learning models to run efficiently on Ampere’s deep learning accelerators. You will work at the intersection of systems software, performance engineering, and AI enablement, helping deliver high-throughput, low-latency inference and a strong foundation for future model and framework support.

What You’ll Achieve:

  • Build and evolve an AI Runtime Libraryfor Ampere accelerators that supports execution, scheduling, and lifecycle management of deep learning workloads across multiple model types and popular frameworks.
  • Own end-to-end acceleration paths, going deep into the full SW/HW stack—including:
    • Inference serving and integration layers
    • Compiler/runtime interfaces and graph/IR execution flows
    • Runtime library architecture (APIs, memory management, operators, execution engines)
    • Communication mechanisms and device/host orchestration
  • Drive HW/SW co-design and optimizationto improve:
    • Throughput (tokens/requests per second)
    • Latency (kernel execution and scheduling efficiency)
    • Memory efficiency (buffering, paging, reuse, caching)
    • Overall compute utilization and scaling behavior
  • Contribute to AI co-processor/accelerator software enablement, partnering closely with hardware and systems teams to ensure runtime and kernel strategies match accelerator capabilities and constraints.
  • Collaborate cross-functionallyto integrate runtime components into Ampere platform stacks, ensuring robust deployment on target environments and consistent performance in production-like workloads.

Ability to operate effectively in a collaborative environment—owning complex components while partnering with compilers, hardware, and platform teams.

About You:

  • BS Computer Science, Computer Engineering, Electrical Engineering, or Software Engineering or related technical field & 8 years of related experience; or MS degree & 6 years; or PhD & 3 years
  • Proven experience developing user-mode drivers and/or runtime libraries for GPUs or deep learning accelerators in Linux or RTOS environments.
  • Strong expertise in C/C++ and systems-level programming (memory, threading, synchronization, performance profiling).
  • Demonstrated background in AI framework enablement, with hands‑on experience in one or more of:
    • PyTorch (operator/runtime integration, graph execution, correctness/performance work)
    • llama.cpp (inference/runtime execution patterns)
    • ONNX (graph handling, interoperability, execution engines)
  • Strong performance engineering skills, including profiling/diagnostics and optimization of execution pipelines, data movement, and compute kernels.

What we’ll offer:

At Ampere we believe in taking care of our employees and providing a competitive total rewards package that includes base pay, cash long-term incentive, and comprehensive benefits.

The full base pay range for this role is between 301,500 PLN and 452,500 PLN.

Benefit highlights include:

  • Premium medical health care, so that you and your family members can feel secure in your health.
  • A generous paid time off policy so that you can embrace a healthy work-life balance.
  • A wide variety of office amenities including nutritious snacks and refreshing drinks, free gym and sauna access to keep you fueled and healthy.
  • Flexible working hours and a remote work policy that includes reimbursement of connectivity costs and equipment to work from home.
  • Sports card fully financed by Ampere.

And there is much more than compensation and benefits. At Ampere, we foster an inclusive culture that empowers our employees to do more and grow more. We are excited to share more about our career opportunities with you through the interview process.

Ampere is an inclusive and equal opportunity employer and welcomes applicants from all backgrounds. All qualified applicants will receive consideration for employment without regard to race, color, national origin, citizenship, religion, age, veteran and/or military status, sex, sexual orientation, gender, gender identity, gender expression, physical or mental disability, or any other basis protected by federal, state or local law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Accelerator Software Principal Engineer – Runtime Library
AI Accelerator Software Principal Engineer – Runtime Library

Ampere • Warszawa

Hybrid
PLN 301,000 - 453,000
Premium medical health care
Generous paid time off
Office amenities: snacks, gym access
+2
AI Accelerator Software Senior Principal Engineer- Framework Integration
AI Accelerator Software Senior Principal Engineer- Framework Integration

Ampere • Warszawa

Hybrid
PLN 416,000 - 625,000
Premium medical health care
Generous paid time off
Office amenities
+2
AI Accelerator Software Senior Principal Engineer- Framework Integration
AI Accelerator Software Senior Principal Engineer- Framework Integration

Ampere Computing • Warszawa

Hybrid
PLN 417,000 - 625,000
Premium health care
Paid time off
Office amenities
+2
AI Accelerator Software Principal Engineer- Framework Integration
AI Accelerator Software Principal Engineer- Framework Integration

Ampere • Warszawa

On-site
PLN 302,000 - 453,000
Premium medical health care
Generous paid time off
Office amenities and snacks
+1
Principal AI Accelerator Runtime Engineer
Principal AI Accelerator Runtime Engineer

Ampere Computing • Warszawa

On-site
PLN 302,000 - 453,000
Premium medical care
Generous PTO
Gym and sauna access
+3
Remote AI Accelerator Framework Engineer - Principal
Remote AI Accelerator Framework Engineer - Principal

Ampere • Warszawa

On-site
PLN 302,000 - 453,000
Premium medical health care
Generous paid time off
Office amenities and snacks
+1
Senior AI Accelerator Software Architect (Framework & Perf) - Remote
Senior AI Accelerator Software Architect (Framework & Perf) - Remote

Ampere Computing • Warszawa

Hybrid
PLN 417,000 - 625,000
Premium health care
Paid time off
Office amenities
+2
Remote AI Accelerator Software Lead — Framework & Performance
Remote AI Accelerator Software Lead — Framework & Performance

Ampere • Warszawa

Hybrid
PLN 416,000 - 625,000
Premium medical health care
Generous paid time off
Office amenities
+2
Senior C++ Software Engineer with CUDA/GPU/TPU
Senior C++ Software Engineer with CUDA/GPU/TPU

EPAM Systems • Poland

Hybrid
PLN 240,000 - 360,000
Hybrid by design
Remote work within Poland
Relocation opportunities
+4
Lead C++ Developer
Lead C++ Developer

EPAM Systems • Poland

Hybrid
PLN 320,000 - 520,000
Hybrid by design
Remote work within Poland
Relocation opportunities
+4