A complete application in a minute — tailored resume and cover letter, ready to send.
Artificial Intelligence Underwriting Company (AIUC) is hiring an engineer to own end-to-end customer engagements. You will scope tests, integrate with customer systems, and run evaluations to validate adherence to the AIUC-1 standard.
You’ll build bespoke integrations when needed, deliver results directly to customers, and collaborate with delivery, engineering, and sales teams to shape future products. This high-ownership role is based in SF with relocation supported where needed.
AIUC builds standards and insurance for frontier AI. Our mission is to Underwrite superintelligence. We exist because risk, not capability, is becoming the binding constraint on AI being useful. History offers lessons for how to build confidence in new technology; electricity burned down houses until insurers funded Underwriters Laboratories (UL) to certify products against their standards. Today, the UL mark is a trustmark on every light bulb in America. We're building that for AI.
AIUC-1, our standard for agent security, serves frontier builders like Cursor, Lovable, ElevenLabs, and Harvey. We've raised $55m from Ribbit, First Harmonic and Nat Friedman to audit frontier AI. Our team comes from METR, Anthropic, McKinsey, and OpenAI. Come join us.
Every AI company that comes to us wants the same thing: proof their agent is safe enough to sell to an enterprise. Getting there means running thousands of evaluations against their live system.
We've now certified the category leaders in every major AI vertical: ElevenLabs (voice), Intercom (customer service), UiPath (workflow automation), Lovable (coding), and more we'll announce soon. No two of those systems look alike: each has its own authentication, its own interface, its own definition of what a failure even looks like. Before we can test any of it, someone has to sit with that customer, understand how their agent actually works, and get it wired into our evaluation system.
We're hiring an engineer to own that work end to end. You'll have your own set of customers within your first weeks, paired with a commercial counterpart on each one.
That last one is why the role exists. We're building toward a product that runs evaluations end to end without an engineer in the loop. This is a high-agency role: you'll shape what we build next as much as you'll deliver what we've built.
If you're planning to start a company one day, this is a front row seat. You'll work directly with our founders, with founders at the companies we certify, and alongside a team full of ex-founders.
The job in practice looks like reading someone's API docs, chasing credentials, building integrations into customers' systems and agents. Then evaluating whether those agents behave, and owning the failure modes the evals surface. The person who thrives here has done this work before. What drives them is trust and safety.
We are in the business of building trust. We take our values seriously and treat them as commitments to each other, held to the same bar as the ones we make to customers.
Underwriting superintelligence is our life's work. If you want it to be yours.