Get more replies from employers
Send a job-specific resume in minutes.
F5 is seeking an AI Inference Engineer to bridge high-performance model development with optimized deployment. You will optimize Large Language Models for inference across GPU-rich data centers and edge devices, focusing on throughput, latency, and accuracy.
You will work on hardware acceleration, scalable infrastructure, and performance monitoring to ensure enterprise-grade reliability and efficient AI capabilities.
At F5, we strive to bring a better digital world to life. Our teams empower organizations across the globe to create, secure, and run applications that enhance how we experience our evolving digital world. We are passionate about cybersecurity, from protecting consumers from fraud to enabling companies to focus on innovation.
Everything we do centers around people. That means we obsess over how to make the lives of our customers, and their customers, better. And it means we prioritize a diverse F5 community where each individual can thrive.
The AI Inference Engineer plays a critical role in the AI lifecycle by bridging the gap between high-performance model development and optimized deployment environments. This position focuses on optimizing Large Language Models (LLMs) for inference, serving diverse environments—from GPU-rich data centers to resource-constrained edge devices with a strong emphasis on maximizing throughput, minimizing latency, and maintaining model accuracy.
This role is pivotal in advancing F5’s AI capabilities, ensuring enterprise-grade reliability by leveraging hardware acceleration, designing scalable infrastructure, and monitoring system performance.