An application made for this job — a tailored resume and cover letter that speak straight to the posting.
DeepInfra is shaping AI inference at scale, and you will own the technical win from first call through production for enterprise deals. You will run benchmarks, tune deployments on cutting-edge hardware, and create reusable assets to accelerate future deals.
This is a high-impact, hands-on role in a fast-growing startup with close collaboration across Sales, Co-founders, and Engineering. We expect strong Python skills, a track record of delivering customer-facing outcomes in ML platforms, and
DeepInfra is building the foundation for companies to use modern AI in production — simply, reliably, and at scale. Our team has deep experience building large systems that serve hundreds of millions of users, and we're bringing that same level of rigor to a rapidly evolving AI inference space. Our mission is to make advanced AI available to people and teams everywhere.
We're an early, tight-knit team where you can influence product direction, try bold ideas, and drive meaningful work forward quickly. If you want to join a fast-growing company at a defining moment, we'd love to talk.
DeepInfra is backed by leading investors including A.Capital, Felicis, 500 Global, Georges Harik, Samsung Next, Supermicro, Upper90, Peak6, SVAngel and Nvidia.
As DeepInfra's enterprise pipeline grows, our customers need a technical partner who can run rigorous evals, defend benchmarks, and speak fluently to both engineering and procurement — someone who can own the technical win from first call through production.
This is a pioneering role. You'll work closely with Sales, our co-founders, and the engineering team on the deals that matter most. You'll own the technical win end to end: running head-to-head bake-offs against leading AI providers, tuning deployments on the latest hardware, and turning what you learn into reusable assets that make every future deal faster to close. As our first FDE, you'll also define what the function looks like as GTM scales.
Three traits define the people who thrive here, and this role leans on all three.
Initiative. We take ownership and step in where we can add value. Whether it's starting something new, improving what exists, or helping move ideas forward, we aim to be proactive and thoughtful in how we contribute.
Drive. We're energized by hard problems. Building AI infrastructure is complex, and we lean into that. We care about doing things well, moving fast, and continuously improving — because solving meaningful challenges is what motivates us.
Grit. Things don't always work on the first try — and that's expected. We stay persistent, adapt quickly, and learn as we go. We take setbacks seriously, but not personally, and use them to get better.
The base pay range for this role is $150,000 – $195,000 per year.