A complete application in a minute — tailored resume and cover letter, ready to send.
Axelera AI is seeking an AI System Engineer to build and maintain the Wingman platform and its agentic AI stack. You will create agent capabilities, execution environments and tool integrations to run AI workloads, validate the platform, and guide improvements.
The role involves deploying models, verifying platform correctness on real hardware, and ensuring reliable operation in diverse environments. Ideal candidates have hands-on experience with AI agents, LLM services and computer vision
Axelera AI https://www.axelera.ai is not your regular deep-tech company. We are creating the next-generation AI platform to support anyone who wants to help advancing humanity and improve the world around us.
In just five years, we have raised a total of $450 million and have built a world-class team of 250+ employees (including 60+ PhDs with more than 40,000 citations), both remotely from 20 different countries and with offices in Belgium, France, Switzerland, Italy, the UK, headquartered at the High Tech Campus in Eindhoven, Netherlands.
We have also launched our Metis™ AI Platform, which achieves a 3-5x increase in efficiency and performance, and have visibility into a strong business pipeline exceeding $100 million.
Our unwavering commitment to innovation has firmly established us as a global industry pioneer.
Are you up for the challenge?
We're looking for an AI System Engineer to help build and maintain Axelera Wingman and Axelera's agentic AI platform. You'll develop the agent capabilities, execution environments and software integrations that make the platform reliable and useful. Deploying models and running AI workloads will help you validate the platform, identify gaps and improve the product.
We value demonstrated ability and judgment over particular degrees or certifications
Software Engineering: Python, Rust or another systems language. APIs, databases and native integrations. TypeScript/React and Tauri experience
Workload Orchestration: Job submission, scheduling, quotas, concurrency. Result retrieval and reproducible environments. SDK and driver compatibility management
Security: Authentication, authorization, least privilege. Access revocation, sandboxing, network security. Protection of users' files and data
Cloud Operations (Google Cloud/AWS): Containers, infrastructure as code, CI/CD. Monitoring, incident investigation. Safe deployments, rollback and recovery
Applied ML & Accelerated Computing: PyTorch/ONNX. Image/video processing, detection, classification, segmentation. Calibration, quantization and hardware-aware optimization
Evaluation & Retrieval: Reproducible benchmarks for model quality, agent behavior and task success. Grounded retrieval and traceable results
We offer a flexible working arrangement, with options to: