Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Tensordyne in Sunnyvale, CA, seeks a Systems Performance Modeling Engineer to build models and tools predicting AI inference workloads on our systems from single accelerators to rack-scale deployments. This hands-on role involves writing simulator code, running experiments, and analyzing discrepancies between predictions and measurements.
You'll collaborate across architecture, silicon, hardware, and software teams to extend models of silicon, interconnects, and fabric, capture workload
Artificial intelligence (AI) is transforming our world. It can perform cognitive functions that previously only humans could do, such as perceiving interactions across different modalities and environments - with the ability to quickly learn and then solve complex problems. Tensordyne is an AI system solution company that builds very high-performance, low-power generative AI inference systems. Our mission, through the creation of custom silicon, hardware and software, is to enable multimodal Generative AI inference acceleration at scale, with safe, sustainable, high-performance systems for our hyperscaler and neocloud data center customers. We are at the leading edge of advancing the latest research and product improvements for generative Al inference solutions that will make Al even more advantageous for compelling new generative AI applications. Tensordyne is a well funded, fast-paced startup company with headquarters in both Sunnyvale, CA, and Munich, Germany. We also have many talented team members working remotely across North America and Europe. We take care of our people and their families with comprehensive benefits, competitive compensation, flexible spending options, and recognition programs, because building category-defining technology starts with a healthy, supported team. Come join us as we shape the future of multimodal generative artificial intelligence!
We are looking for a Systems Performance Modeling Engineer to build the models and tools that predict how generative AI inference workloads perform on Tensordyne systems, from a single accelerator up through rack, pod, and cluster scale. This is a hands-on engineering role for someone who likes writing simulator code, running experiments, and digging into why a prediction and a measurement don't match.
Working closely with our architects and the silicon, hardware, networking, and software teams, you'll capture workload behavior, extend simulation and analytical models of our silicon, interconnect, and multi-hop fabrics, and validate them against real hardware. You'll be comfortable moving across the stack, from the model graph through collectives to the network fabric, to track down where performance is going.
Tensordyne is an equal opportunity employer. We believe that a diverse team is better at tackling complex problems and coming up with innovative solutions. All qualified applicants will receive consideration for employment without regard to age, color, gender identity or expression, marital status, national origin, disability, protected veteran status, race, religion, pregnancy, sexual orientation, or any other characteristic protected by applicable laws, regulations and ordinances.