Turn this role into an interview — a resume and cover letter built around what this employer wants.
Oho Group is building an AI software stack and compiler infrastructure in San Francisco. This role sits at the crossroads of compiler optimization and hardware mapping, transforming models and graphs into efficient programs for a new accelerator architecture.
You will own optimization passes, contribute to MLIR-based pipelines, and work with architecture, code-generation and runtime teams to push performance of AI workloads on modern GPUs and accelerators.
We’re working with an AI software company building a new compiler and software stack for high-performance accelerated computing.
This role sits in the middle of the compiler, transforming models and computational graphs into efficient programs for a new accelerator architecture. You’ll own optimizations that materially affect workload performance rather than maintaining an established toolchain.
Experience with PyTorch compilation, Triton, GPUs, AI accelerators, memory optimization or hardware/software co-design would be particularly valuable.
This is an opportunity to shape a compiler while both the software stack and target architecture are still evolving.