Get more replies from employers
Send a job-specific resume in minutes.
uRun is seeking a Site Reliability Engineer to own the reliability culture from scratch. As a founding hire, you'll define the observability stack and incident response processes, working closely with infrastructure engineers to ensure uptime and efficiency.
We're building cutting-edge infrastructure for interactive AI, and this role is pivotal in shaping our foundation. A strong background in SRE, cloud infrastructure, and incident management is essential.
Enjoy competitive salary and equity, along with comprehensive health coverage, flexible spending accounts, and the best tools available.
Most AI infrastructure is built for batch: send a query, wait, get a response, reset. Powerful, but transactional. AI is becoming interactive — sessions that hold state, models that stay alive between turns, generation that responds as it runs — and the infrastructure to deliver that at scale doesn't really exist yet.
The bottleneck isn't the models anymore. It's the infrastructure underneath them.
uRun is the inference cloud for interactive AI: the compute layer that makes real-time, stateful inference possible at scale. We came out of stealth in April 2026, are backed by top-tier investors, and are founded by Keegan McCallum, who scaled inference infrastructure for some of the most demanding generative AI workloads in production.
We're an infrastructure company. We build the layer that model labs, builders, and research teams ship on top of.
Reliability at uRun isn't a feature — it's the product. When model labs and production teams build on top of our inference platform, they are trusting us with their uptime, their latency, and their users. As our Site Reliability Engineer, you will own that trust end-to-end.
This is a founding SRE hire. You will define the reliability culture from scratch: the observability stack, the incident response playbooks, the SLOs, and the on-call process. You will work directly with infrastructure and platform engineers to close the gap between what we ship and what stays up.
Competitive salary and meaningful equity in an early‑stage AI infrastructure company. The band above is our target; for an exceptional candidate we’ll go higher. Equity is real — you’re early, and the grant reflects that.
We build the stage, not the show. We’re an infrastructure company, a developer‑tools company, and a production partner for model labs, and focus is a deliberate choice we’ve made and hold to.
Day‑to‑day, that means a small team, a high bar, and real ownership. You won’t wait for permission or inherit a backlog of someone else’s decisions; in a founding security role, the function is what you make it.
It also means ambiguity: priorities shift, not everything is documented, and you’ll often be the person who decides what "secure enough, for now" means. That suits some people and not others, and we’d rather you know that before you apply.