Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Oho Group in San Francisco is seeking a seasoned compiler engineer to own optimizations that transform models and graphs into efficient programs for a new accelerator architecture. You will shape the compiler mid-stack, focusing on performance-critical paths rather than maintaining an existing toolchain.
You’ll work across MLIR-based pipelines, optimize data movement, parallelism and hardware utilization, and collaborate with architecture, code-generation and runtime engineers to map ML
We’re working with an AI software company building a new compiler and software stack for high-performance accelerated computing.
This role sits in the middle of the compiler, transforming models and computational graphs into efficient programs for a new accelerator architecture. You’ll own optimizations that materially affect workload performance rather than maintaining an established toolchain.
What you’ll work on
What we’re looking for
Experience with PyTorch compilation, Triton, GPUs, AI accelerators, memory optimization or hardware/software co-design would be particularly valuable.
This is an opportunity to shape a compiler while both the software stack and target architecture are still evolving.