Remote AI Inference Engineer: Kernel & Edge Optimization

Lever, Inc.

Österreich

Remote

EUR 90.000 - 130.000

Vollzeit

Vor 7 Tagen
Sei unter den ersten Bewerbenden
Bewerbungsgenerator

Mach aus dieser Rolle ein Bewerbungsgespräch — ein Lebenslauf und ein Anschreiben, die genau auf das zugeschnitten sind, was dieser Arbeitgeber sucht.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Remote-first team
International collaboration
Cutting-edge AI research
Performance-critical infrastructure
Multimodal AI projects

Zusammenfassung

Lever, Inc. is seeking an AI Research Engineer (Kernel & Inference Optimization) based in Austria.

You will work at the intersection of AI research, systems engineering, and high-performance model inference, focusing on model-serving architectures across hardware environments including edge devices. The role blends hands-on research with low-level engineering to translate findings into measurable performance improvements.

Qualifikationen

  • PhD in NLP, ML, or related field with strong AI research track record.
  • Expertise in Metal Shading Language (MSL) with ability to write custom compute shaders from scratch.
  • Experience with low-level kernel optimization and inference optimization on mobile or resource-constrained devices.
  • Proven ability to deliver measurable improvements in inference latency, throughput, and memory footprint.
  • Knowledge of distributed inference techniques: tensor, pipeline, and expert parallelism.
  • Strong English communication and collaboration with distributed teams.

Aufgaben

  • Design and deploy model-serving architectures optimized for high throughput and low latency.
  • Develop inference pipelines for mobile and edge platforms.
  • Set performance targets for latency, throughput, memory, and reliability.
  • Build benchmarks and track latency, throughput, memory, and error rates.
  • Create datasets and simulation scenarios for evaluating model performance under constraints.
  • Identify bottlenecks and implement system-level optimizations.
  • Develop GPU kernels and compute shaders for mobile hardware (MSL).
  • Apply pruning, quantization, Flash Attention, KV caching, and speculative decoding.
  • Design distributed inference using tensor, pipeline, and expert parallelism.

Kenntnisse

Metal Shading Language
Kernel optimization
Inference optimization
GPU kernels
Edge/mobile deployment
Distributed inference
Benchmarking
Communication in English

Ausbildung

PhD in NLP / ML
BSc in Computer Science

Tools

MSL tooling

Jobbeschreibung

Lever, Inc. is seeking an AI Research Engineer (Kernel & Inference Optimization) based in Austria.

You will work at the intersection of AI research, systems engineering, and high-performance model inference, focusing on model-serving architectures across hardware environments including edge devices. The role blends hands-on research with low-level engineering to translate findings into measurable performance improvements.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

AI Research Engineer (Kernel & Inference Optimization)
AI Research Engineer (Kernel & Inference Optimization)

Lever, Inc. • Österreich

Remote
EUR 90.000 - 140.000
Remote-first team
International collaboration
Cutting-edge AI research
+2
Edge Linux Engineer for AI-Ready Platforms
Edge Linux Engineer for AI-Ready Platforms

SKIDATA • Klagenfurt am Wörthersee

Hybrid
EUR 29.000 - 35.000
Meal allowance
Job bike leasing program
Transport subsidy
+2
Senior AI Inference Engineer - High-Performance GPU Systems
Senior AI Inference Engineer - High-Performance GPU Systems

NVIDIA • Lavamünd

Vor Ort
EUR 120.000 - 160.000
Senior AI Inference Architect — Multi-Node GPU Scale
Senior AI Inference Architect — Multi-Node GPU Scale

NVIDIA • Lavamünd

Vor Ort
EUR 110.000 - 170.000
Staff AI Engineer: Generative AI, LLM & RAG Solutions
Staff AI Engineer: Generative AI, LLM & RAG Solutions

Infineon Technologies AG • Klagenfurt am Wörthersee

Vor Ort
EUR 60.000 - 66.000
AI/ML Specialist — Edge, Cloud & Embedded Solutions
AI/ML Specialist — Edge, Cloud & Embedded Solutions

Palfinger Ag • Lengfelden

Hybrid
EUR 48.000 - 59.000
Flexible working hours; remote option
International career opportunities
PALfit corporate health management
+3
Edge Platform Engineer: Linux & AI-Ready Systems
Edge Platform Engineer: Linux & AI-Ready Systems

Traka (Assa Abloy) • Klagenfurt am Wörthersee

Hybrid
EUR 29.000 - 35.000
Flexible hybrid work model
Meal allowance
Job bike leasing program
+5
Staff AI Engineer - Generative AI & LLM Solutions
Staff AI Engineer - Generative AI & LLM Solutions

Infineon Technologies • Klagenfurt am Wörthersee

Vor Ort
EUR 62.000 - 75.000
Senior AI Platform Engineer (MLOps)
Senior AI Platform Engineer (MLOps)

RetInSight GmbH • Wien

Hybrid
EUR 58.000 - 71.000
Hybrid work model
Central Vienna office
Public transport support
+2
Edge Platform Engineer - Linux & AI-Ready (Hybrid)
Edge Platform Engineer - Linux & AI-Ready (Hybrid)

Record UK Ltd • Klagenfurt am Wörthersee

Hybrid
EUR 29.000 - 35.000
Flexible hybrid work model
Meal allowance
Job bike leasing program
+3