A complete application in a minute — tailored resume and cover letter, ready to send.
Huawei is seeking an AI Inference & Compression Engineer in Singapore to advance LLM inference acceleration, develop advanced video codecs, and optimize AI-based media coding. The role includes deploying optimized models and integrating compression techniques into deployment frameworks such as vLLM.
The candidate will work on performance and quality evaluation, including PSNR, VMAF, perplexity, latency, and throughput, ensuring robust, scalable AI/video systems.
Huawei is seeking an AI Inference & Compression Engineer in Singapore to advance LLM inference acceleration, develop advanced video codecs, and optimize AI-based media coding. The role includes deploying optimized models and integrating compression techniques into deployment frameworks such as vLLM.
The candidate will work on performance and quality evaluation, including PSNR, VMAF, perplexity, latency, and throughput, ensuring robust, scalable AI/video systems.