Turn this role into an interview — a resume and cover letter built around what this employer wants.
d-Matrix Frontier Group in Santa Clara, CA is seeking end-to-end inference engineers to drive novel ideas to deployed, optimized systems across the inference stack. You will work from kernel-level optimization to distributed orchestration and high-level serving APIs, shaping the future of AI silicon usage.
Ideal candidates have deep experience with LLM inference, open-source frameworks, and heterogeneous hardware deployments, delivering POCs to customers and contributing to open-source projects.
d-Matrix Frontier Group in Santa Clara, CA is seeking end-to-end inference engineers to drive novel ideas to deployed, optimized systems across the inference stack. You will work from kernel-level optimization to distributed orchestration and high-level serving APIs, shaping the future of AI silicon usage.
Ideal candidates have deep experience with LLM inference, open-source frameworks, and heterogeneous hardware deployments, delivering POCs to customers and contributing to open-source projects.