ML Runtime Engineer — Edge Inference & Equity

ATX Venture Partners

Occitanie

Presencial

EUR 85 000 - 110 000

Tempo integral

14 dias+
Gerador de candidaturas

Não envies um currículo genérico — gera um currículo e uma carta de apresentação adaptados a esta função específica.

Ultrapassa os filtros ATS

Vantagens oferecidas por esta oferta de emprego

Equity
Paid time off
Health insurance
Lunch vouchers
Cross-office travel
Off-sites
Relocation assistance

Resumo da oferta

ATX Venture Partners is seeking an experienced engineer to own model compilation for accelerator targets, turning ONNX models into optimized engines while respecting on-device power and thermal budgets.

The role focuses on building the ground-side compile and delivery path, performance benchmarking, and contributing to the Space Inference Engine and SDK. You will coordinate with the OBSW/Runtime team to ensure robust, scalable model builds.

Qualificações

  • Strong C++ and Python programming skills.
  • Experience with model compilation and graph compilers (TensorRT).
  • Experience with ONNX Runtime and performance optimization.
  • Experience with embedded/edge GPU deployment (NVIDIA Jetson).
  • Experience benchmarking and profiling with perf tools.
  • CI/CD experience including hardware-in-the-loop testing.

Responsabilidades

  • Own model compilation: turn partner ONNX models into optimized engines per accelerator.
  • Build the ground-side compile & delivery path; define how a mismatched engine is detected/rejected on-node.
  • Build performance tooling: benchmarking, profiling, operator-coverage matrices, and budget validation.
  • Contribute to Space Inference Engine (execution providers) and the SDK (model build & packaging). Coordination of OBSW / Runtime team.

Conhecimentos

C++ and Python
TensorRT
ONNX Runtime & perf optimization
Embedded / edge GPU deployment
Benchmarking & profiling
CI/CD with hardware-in-the-loop

Ferramentas

Hailo SDK
Meson build
Yocto / minimal Debian images

Descrição da oferta de emprego

ATX Venture Partners is seeking an experienced engineer to own model compilation for accelerator targets, turning ONNX models into optimized engines while respecting on-device power and thermal budgets.

The role focuses on building the ground-side compile and delivery path, performance benchmarking, and contributing to the Space Inference Engine and SDK. You will coordinate with the OBSW/Runtime team to ensure robust, scalable model builds.

Obtém a tua avaliação gratuita e confidencial do currículo.

ou arrasta e larga o ficheiro aqui.

Similar jobs

Ofertas semelhantes que vale a pena comparar

ML Runtime Engineer
ML Runtime Engineer

ATX Venture Partners • Occitanie

Presencial
EUR 85 000 - 110 000
Equity
Paid time off
Health insurance
+4
ML Inference Engineer | Accelerate AI on Custom Accelerator
ML Inference Engineer | Accelerate AI on Custom Accelerator

Arago • Paris

Presencial
EUR 90 000 - 130 000
Stock options
Healthcare coverage
Pension contributions
+2
Inference Systems Engineer: Scale & Performance
Inference Systems Engineer: Scale & Performance

Mistral • Paris

Híbrido
EUR 120 000 - 180 000
Healthcare coverage
Relocation support
Wellness programs
Senior AI Engineer
Senior AI Engineer

Jobtailor • França

Presencial
EUR 75 000 - 110 000
On-Device AI Engineer for Real-Time Gaming
On-Device AI Engineer for Real-Time Gaming

Jobtailor • França

Presencial
EUR 75 000 - 110 000
ML Inference Systems Engineer — High-Performance Accelerator
ML Inference Systems Engineer — High-Performance Accelerator

Arago Inc. • Paris

Presencial
EUR 110 000 - 170 000
Stock options
Health insurance
Pension contributions
+1
Hybrid ML Engineer: Hardware-Aware Model Optimization
Hybrid ML Engineer: Hardware-Aware Model Optimization

Aimlroles • Paris

Híbrido
EUR 75 000 - 90 000
Competitive salary
Hybrid work in Paris
Meal & wellness benefits
Runtime Architect for Co-Designed AI Accelerators
Runtime Architect for Co-Designed AI Accelerators

HiPEAC • Bordeaux

Presencial
EUR 45 000 - 70 000
AI Accelerator Architect & Performance Modeler
AI Accelerator Architect & Performance Modeler

Arago Inc. • Paris

Presencial
EUR 150 000 - 210 000
Stock option plan
Healthcare coverage
Pension contributions
+1
AI Accelerator Architecture Modeling Engineer
AI Accelerator Architecture Modeling Engineer

arago • Paris

Presencial
EUR 90 000 - 120 000
Stock options
Healthcare coverage
Pension contributions
+3