Stand out for this role — generate a tailored resume and cover letter in about a minute.
Lupitor is building the infrastructure layer for enterprise voice AI, delivering on-prem and air-gapped deployments for regulated environments. You will design, implement, and iterate the systems powering live customer conversations at scale, taking ideas from concept to production and owning refinement end to end.
You will optimize the inference stack, focusing on latency, throughput, and cost, and ensure portability and operability across diverse customer environments while maintaining high
Lupitor is the infrastructure layer for enterprise voice AI. We build agents that handle real customer conversations in production, running on the customer's own infrastructure instead of the cloud. The bet underneath: open source and data sovereignty out-compete cloud AI long term, and voice is where that bet lands first. To make that work we build the inference layer ourselves - serving and caching engines for TTS and speech-to-speech - because owning the stack is what makes on-prem deployment actually fast enough to use.
You will design, implement, and iterate on the systems that power live customer conversations at scale - from call orchestration and telephony to the inference path that makes real-time voice possible. You will take ideas from concept to production, owning refinement end to end.
You will work on our serving and caching engines for TTS and speech-to-speech - optimizing latency, throughput, and cost in the places where the product actually wins or loses. This means profiling, benchmarking, and making real-time systems faster under real load.
You will ship our platform into on-prem and air-gapped customer infrastructure - banks and governments with their own security, network, and operational constraints. You will think about portability, operability, and failure modes, not just happy paths.
You will balance speed and quality, shipping fast while maintaining a strong standard for reliability, performance, and code health. Voice systems fail loudly - yours shouldn't.
Experience using our stack is not necessary for this role, but you are expected to have worked on enterprise products at scale with a modern frontend and backend. That said, we use Go for our servers, TypeScript for our web apps (Next.js) and backend, and AWS for hosting. We use Convex on top of PostgreSQL for DB and manage our infrastructure with Terraform.