Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Oho Group in San Francisco, CA is seeking a Principal Performance Modeling Architect for AI systems to lead architecture-informed performance modeling before hardware realization. You will build analytical and simulation models to guide decisions across compute, memory, and interconnect, influencing silicon and software direction.
The role emphasizes strong Python/C++ skills, deep architectural knowledge, and collaboration with hardware, compiler, and runtime teams to translate model results
We’re working with a semiconductor company designing a new programmable compute architecture for demanding AI workloads.
They’re looking for a senior performance-modeling engineer to create the tools that guide architectural decisions before hardware exists. Your work will determine where performance is gained or lost across compute, memory, interconnect and complete AI systems.
Experience with GPUs, AI accelerators, LLM inference, cluster modelling, roofline analysis, NoCs or cost-per-token analysis would be especially relevant.
This is an architecture-shaping role where performance models influence both the silicon and the software designed around it.