ML Runtime Engineer

ATX Venture Partners

Occitanie

Sur place

EUR 85 000 - 110 000

Plein temps

14 jours+

Recevez plus de réponses des employeurs

Envoyez un CV adapté au poste en quelques minutes.

Avantages offerts par ce poste

Equity
Paid time off
Health insurance
Lunch vouchers
Cross-office travel
Off-sites
Relocation assistance

Résumé du poste

ATX Venture Partners is seeking an experienced engineer to own model compilation for accelerator targets, turning ONNX models into optimized engines while respecting on-device power and thermal budgets.

The role focuses on building the ground-side compile and delivery path, performance benchmarking, and contributing to the Space Inference Engine and SDK. You will coordinate with the OBSW/Runtime team to ensure robust, scalable model builds.

Qualifications

  • Strong C++ and Python programming skills.
  • Experience with model compilation and graph compilers (TensorRT).
  • Experience with ONNX Runtime and performance optimization.
  • Experience with embedded/edge GPU deployment (NVIDIA Jetson).
  • Experience benchmarking and profiling with perf tools.
  • CI/CD experience including hardware-in-the-loop testing.

Responsabilités

  • Own model compilation: turn partner ONNX models into optimized engines per accelerator.
  • Build the ground-side compile & delivery path; define how a mismatched engine is detected/rejected on-node.
  • Build performance tooling: benchmarking, profiling, operator-coverage matrices, and budget validation.
  • Contribute to Space Inference Engine (execution providers) and the SDK (model build & packaging). Coordination of OBSW / Runtime team.

Connaissances

C++ and Python
TensorRT
ONNX Runtime & perf optimization
Embedded / edge GPU deployment
Benchmarking & profiling
CI/CD with hardware-in-the-loop

Outils

Hailo SDK
Meson build
Yocto / minimal Debian images

Description du poste

About this role:
  • Own model compilation: turn partner ONNX models into optimized engines per accelerator (TensorRT, Hailo, AMD/ROCm), within power & thermal budgets
  • Build the ground-side compile & delivery path; define how a mismatched engine is detected/rejected on-node
  • Build performance tooling: benchmarking, profiling, operator-coverage matrices, and budget validation
  • Contribute to Space Inference Engine (execution providers) and the SDK (model build & packaging); coordinate the OBSW / Runtime team
Must Haves:
  • Strong C++ and Python
  • Model compilation: TensorRT (and/or equivalent graph compilers)
  • ONNX Runtime, quantization & inference perf optimization
  • Embedded / edge GPU deployment (NVIDIA Jetson)
  • Benchmarking & profiling / perf tooling
  • CI/CD incl. hardware-in-the-loop
Nice to Haves:
  • Hailo SDK and/or AMD ROCm compilation (iX10)
  • Writing custom kernels / operator plugins
  • Remote sensing / large-image handling
  • RF / IQ signal data exposure
  • Meson build, Yocto / minimal Debian images
  • Model-weight protection / secure execution
Some of Our Awesome Benefits:
  • Equity, we want you to have an active role in our success
  • Up to 35 days of Paid Time Off (vacations & RTT ) and flexible working hours, we want you to be at your best
  • Health and life insurance, we care about your health
  • Lunch Vouchers, because we love food!
  • Cross-office travel opportunities between San Francisco, Colorado, and Toulouse to learn from our differences
  • Company and team off-sites and many other events to work & celebrate together
  • Relocation assistance to Toulouse when applicable

Research shows that while men apply to jobs where they meet an average of 60% of the criteria, women and other underrepresented people tend to only apply when they meet 100% of the qualifications. At Loft, we value respectful debate and people who aren’t afraid to challenge assumptions. We strongly encourage you to apply, even if you don’t check all the boxes.

Obtenez votre examen gratuit et confidentiel de votre CV.
ou faites glisser et déposez votre fichier ici.
Similar jobs

Postes similaires à comparer

ML Runtime Engineer
ML Runtime Engineer

Loftorbital • Toulouse

Sur place
EUR 55 000 - 75 000
Equity
Flexible working hours
Health and life insurance
+3
Onboard Software Engineer - Deployment & Runtime Environments
Onboard Software Engineer - Deployment & Runtime Environments

Loft Orbital • Toulouse

Sur place
EUR 75 000 - 110 000
Equity
Up to 35 days PTO (vacations & RTT)
Health and life insurance
+3
Onboard Software Engineer - Deployment & Runtime Environments
Onboard Software Engineer - Deployment & Runtime Environments

ATX Venture Partners • Occitanie

Sur place
EUR 70 000 - 90 000
Equity
35 days PTO
Health insurance
+2
Edge AI Runtime Engineer: ONNX/TensorRT On-Device
Edge AI Runtime Engineer: ONNX/TensorRT On-Device

Loftorbital • Toulouse

Sur place
EUR 55 000 - 75 000
Equity
Flexible working hours
Health and life insurance
+3
Frontend Engineer
Frontend Engineer

ATX Venture Partners • Occitanie

Sur place
EUR 55 000 - 75 000
Equity
35 days PTO
Health and life insurance
+4
MLOps Engineer
MLOps Engineer

White Circle • Paris

Sur place
EUR 65 000 - 85 000
Paid time off
Comprehensive medical insurance
Team off-sites twice a year
Site Reliability Engineer (Network)
Site Reliability Engineer (Network)

Cerebras • Occitanie

Sur place
EUR 90 000 - 130 000
Equity
Up to 35 days PTO and flexible hours
Health and life insurance
+2
AI Runtime / Low-level Software Engineer
AI Runtime / Low-level Software Engineer

Kalray • France

Sur place
EUR 70 000 - 110 000
RSU / stock options
Paid leave RTT
Mobility incentives
+2
Frontend Engineer
Frontend Engineer

Loft Orbital • Toulouse

Sur place
EUR 60 000 - 90 000
Equity
35 days PTO
Health insurance
+4
Sales Systems Engineer
Sales Systems Engineer

Alumni Ventures • France

Sur place
EUR 50 000 - 70 000
Equity participation
Up to 35 days Paid Time Off
Health and life insurance
+3