A complete application in a minute — tailored resume and cover letter, ready to send.
Hoonify is seeking an AI Infrastructure Engineer to design, build, and operate production-grade LLM inference infrastructure. You will optimize GPU-backed workloads, implement robust services, and collaborate with senior engineers to deliver scalable, observable systems.
You will work on multi-GPU, multi-node deployments, profiling across vLLM, SGLang, and TensorRT-LLM, and contribute to runbooks and design notes for the team.
Hoonify is seeking an AI Infrastructure Engineer to design, build, and operate production-grade LLM inference infrastructure. You will optimize GPU-backed workloads, implement robust services, and collaborate with senior engineers to deliver scalable, observable systems.
You will work on multi-GPU, multi-node deployments, profiling across vLLM, SGLang, and TensorRT-LLM, and contribute to runbooks and design notes for the team.