A leading AI infrastructure company in San Francisco is looking for a Senior Systems Performance Engineer to lead the hardware evaluation and scaling of their AI infrastructure. This role involves defining the performance roadmap for their next-generation cloud, focusing on optimizing system efficiency. Ideal candidates have over 5 years of experience with large-scale GPU systems, strong programming skills in Python and C++, and a deep understanding of hardware architectures. The compensation ranges from $172,500 to $210,000 annually.
Qualifications
5+ years experience in end-to-end hardware evaluation, reliability, and scaling of AI infrastructure.
Experience with large-scale GPU infrastructure optimization.
Knowledge of performance modeling for secure environments.
Responsibilities
Lead the evaluation and establishment of New Product Introduction across hardware architectures.
Conduct deep-dive performance evaluations and workload characterizations.
Design and implement 0-to-1 performance methodologies for scaling.
Skills
Expert-level proficiency in Python
Expert-level proficiency in C++
Proven experience in building and optimizing AI application systems
Deep knowledge of x86 and ARM architectures
Ability to write and debug ARMv8 assembly
Tools
Lauterbach Trace32
ARM DS-5
Job description
A leading AI infrastructure company in San Francisco is looking for a Senior Systems Performance Engineer to lead the hardware evaluation and scaling of their AI infrastructure. This role involves defining the performance roadmap for their next-generation cloud, focusing on optimizing system efficiency. Ideal candidates have over 5 years of experience with large-scale GPU systems, strong programming skills in Python and C++, and a deep understanding of hardware architectures. The compensation ranges from $172,500 to $210,000 annually.