Get more replies from employers
Send a job-specific resume in minutes.
Shields Group Search is partnering with a privacy‑focused consumer AI company to hire an Engineering Lead for Inference Optimization. This role blends hands‑on work with people leadership to shape the company’s inference stack and scale GPU performance.
You’ll drive latency improvements, build benchmarks, and guide a small team of engineers. Strong GPU, Python (plus Rust/Go) and LLM inference experience are expected, with equity and crypto token compensation alongside a base salary of
Shields Group Search is working with a leading consumer AI company built on the principles of privacy, free speech, and user sovereignty to hire an Engineering Lead, Inference Optimization.
Our client is the world’s leading consumer AI company built on principles of privacy, free speech, and user sovereignty.
They’re building the Port City of AI, in which millions of individuals, third-party apps, and AI agents gather, interact, and access sophisticated AI resources on a private and permissive foundation.
Their mission is to make artificial intelligence approachable and useful in everyday work—bridging the gap between cutting-edge research and practical, real-world impact.
They’re a fast-moving startup where every team member is expected to make a clear impact. Their culture is rooted in curiosity, ownership, ethical principle, philosophy, and collaboration—whether they’re designing better AI workflows, supporting their growing community, or shaping the future of human-AI interaction.
Joining the company means joining a team of unorthodox builders who believe in moving quickly, delivering a beautiful, mass-market, highly useful consumer product that doesn’t spy on people or censor their ideas and questions, and maintaining an edge in the rapidly evolving world of agentic machine intelligence.
If you’re energized by big ideas, entrepreneurial spirit, individual empowerment, and the opportunity to help shape a fast-growing company in the world’s hottest industry from the ground up, you’ll feel right at home here.
Our client is the only AI platform that runs inference with zero data retention and zero training on user inputs.
This is an opportunity to be on the bleeding edge of privacy-focused AI with a unique and dedicated team of high-agency individuals alongside you.
This role requires both hands-on work as an individual contributor as well as the management of a small team. You will play a pivotal role, shaping the company’s overarching technical strategy and assembling an exceptional team to deliver peak inference performance at massive scale.
The base annual salary for this position ranges from $270,000–$330,000 USD, with equity and crypto token compensation included, and reports to the Head of Engineering.