A complete application in a minute — tailored resume and cover letter, ready to send.
Hewlett Packard Enterprise is seeking a Principal Software Engineer to lead the LLM inference runtime within the AI Essentials platform. You will architect engine integration, batching strategies, KV cache reuse, and distributed execution on customer-owned hardware, while coordinating with Kubernetes orchestration and performance teams.
The role emphasizes deep expertise in LLM runtimes, multi-GPU scaling, and deployment at scale.
Hewlett Packard Enterprise is seeking a Principal Software Engineer to lead the LLM inference runtime within the AI Essentials platform. You will architect engine integration, batching strategies, KV cache reuse, and distributed execution on customer-owned hardware, while coordinating with Kubernetes orchestration and performance teams.
The role emphasizes deep expertise in LLM runtimes, multi-GPU scaling, and deployment at scale.