Principal AI Systems Architect — Inference & Hardware
Conductor
San Jose (CA)
On-site
USD 219,000 - 351,000
Full time
14 days+
Get more replies from employers
Send a job-specific resume in minutes.
Start fresh or import an existing resume
Benefits offered by this job
4+ weeks of paid time off
Medical/Dental/Vision/401k
Flexible work environment
Onsite gym and café
Charitable giving match
Job summary
A technology company in San Jose is seeking a Principal Engineer, AI Serving Framework Architect. This role involves leading research teams and leveraging expertise in AI workloads. Ideal candidates will have a PhD and over 15 years of experience in large-scale computing, taking the lead on projects around AI inference performance. An inclusive workplace offers a competitive salary and extensive benefits, emphasizing professional and personal growth.
Qualifications
15+ years of experience in AI Serving Framework for large-scale computing.
Led project to build and optimize a Large Language Model (LLM) inference software stack.
In-depth understanding of inference engines such as vLLM.
Responsibilities
Lead research teams and propose technical direction.
Investigate dynamic scheduling methodologies for AI inference performance.
Propose software design for optimization algorithms on open-source platforms.
Skills
PyTorch
Python
C++
Collaboration
Communication
Education
PhD in Computer Science or a related field
Job description
A technology company in San Jose is seeking a Principal Engineer, AI Serving Framework Architect. This role involves leading research teams and leveraging expertise in AI workloads. Ideal candidates will have a PhD and over 15 years of experience in large-scale computing, taking the lead on projects around AI inference performance. An inclusive workplace offers a competitive salary and extensive benefits, emphasizing professional and personal growth.