A pioneering AI hardware company in Sunnyvale is looking for an engineering leader to oversee the Inference Service Platform. You will guide a team in scaling LLM inference and architecting low latency systems. Candidates should have substantial experience in distributed systems, strong technical leadership, and a history of optimizing performance. The position emphasizes collaboration across teams to deliver enterprise solutions and requires a hands-on approach to technology and team mentorship.
Qualifications
6+ years in high-scale software engineering, with 3+ years leading distributed systems.
Proven track record scaling LLM inference with optimizations.
Deep experience with orchestration and large clusters.
Responsibilities
Own the technical vision for Cerebras Inference Platform.
Lead development of distributed inference systems.
A pioneering AI hardware company in Sunnyvale is looking for an engineering leader to oversee the Inference Service Platform. You will guide a team in scaling LLM inference and architecting low latency systems. Candidates should have substantial experience in distributed systems, strong technical leadership, and a history of optimizing performance. The position emphasizes collaboration across teams to deliver enterprise solutions and requires a hands-on approach to technology and team mentorship.