Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Huawei is seeking an AI Inference & Compression Engineer in Singapore to advance LLM inference acceleration, develop advanced video codecs, and optimize AI-based media coding. The role includes deploying optimized models and integrating compression techniques into deployment frameworks such as vLLM.
The candidate will work on performance and quality evaluation, including PSNR, VMAF, perplexity, latency, and throughput, ensuring robust, scalable AI/video systems.
On behalf of Huawei, a world-renowned information and communication technology company, we are seeking passionate and talented individuals to join our team as AI Inference & Compression Engineer.