Get more replies from employers
Send a job-specific resume in minutes.
uRun is seeking a ML Infrastructure and Platform Engineer to lead the architecture and scaling of its GPU compute platform as a founding technical hire. You will have end-to-end ownership across the infrastructure stack, working directly with the founding team to define systems.
The role requires deep expertise in distributed systems and the ability to ensure high availability and low-latency inference. Competitive salary, equity, and top-tier benefits such as health coverage and a MacBook Pro are offered.
Most AI infrastructure is built for batch: send a query, wait, get a response, reset. Powerful, but transactional. AI is becoming interactive — sessions that hold state, models that stay alive between turns, generation that responds as it runs — and the infrastructure to deliver that at scale doesn't really exist yet.
The bottleneck isn't the models anymore. It's the infrastructure underneath them.
uRun is the inference cloud for interactive AI: the compute layer that makes real-time, stateful inference possible at scale. We came out of stealth in April 2026, are backed by top-tier investors, and are founded by Keegan McCallum, who scaled inference infrastructure for some of the most demanding generative AI workloads in production.
We're an infrastructure company. We build the layer that model labs, builders, and research teams ship on top of.
We are building the next generation of AI inference infrastructure. As our ML Infrastructure and Platform Engineer, you will own the architecture and scaling of our GPU compute platform from the ground up.
This is a founding technical hire with end-to-end ownership across the full infrastructure stack, from bare metal to model serving. You will work directly with the founding team and define how we build.
Competitive salary and meaningful equity in an early-stage AI infrastructure company. The band above is our target; for an exceptional candidate we'll go higher. Equity is real — you're early, and the grant reflects that.
We build the stage, not the show. We're an infrastructure company, a developer-tools company, and a production partner for model labs — and focus is a deliberate choice we've made and hold to.
Day-to-day, that means a small team, a high bar, and real ownership. You won't wait for permission or inherit a backlog of someone else's decisions. In a founding infrastructure role, the function is what you make it.
It also means ambiguity: priorities shift, not everything is documented, and you'll often be the person who decides what "good enough for now" means. That suits some people and not others, and we'd rather you know that before you apply.