A leading semiconductor company is seeking a Principal AI Performance Engineer in San Jose, CA. The ideal candidate will drive performance optimization on AMD GPUs, lead a technical team, and engage with customers to present findings and recommendations. Candidates should have 7+ years of experience in GPU computing or AI systems, deep knowledge of performance diagnostics, and proficiency in Python and C++. This role requires a strong academic background and the ability to leverage AI tools for performance engineering.
Qualifications
7+ years of experience in software development, especially in GPU computing or AI systems.
Hands-on experience with AI serving frameworks and their internals.
Ability to understand and optimize GPU kernel performance.
Responsibilities
Drive end-to-end performance optimization for AI models.
Profile and diagnose performance bottlenecks across multiple systems.
Lead technical engagements with clients, presenting findings and recommendations.
Skills
AI fluency
Profiling tools expertise
Python proficiency
C++ proficiency
Technical leadership
Deep understanding of GPU computing
Experience with AI frameworks
Education
Bachelor’s, Master’s, or PhD in relevant field
Tools
TensorRT
CUDA
Triton
vLLM
Job description
A leading semiconductor company is seeking a Principal AI Performance Engineer in San Jose, CA. The ideal candidate will drive performance optimization on AMD GPUs, lead a technical team, and engage with customers to present findings and recommendations. Candidates should have 7+ years of experience in GPU computing or AI systems, deep knowledge of performance diagnostics, and proficiency in Python and C++. This role requires a strong academic background and the ability to leverage AI tools for performance engineering.