Transforma esta oferta en una entrevista: un currículum y una carta de presentación creados pensando en lo que quiere el empleador.
Wave Group seeks an experienced backend engineer to build, operate and improve the reliability of autonomous AI agents for enterprise clients. You will own the observability and evaluation layer, including tracing, logging, cost tracking, and regression testing across model changes.
You’ll work end-to-end across the stack: production-grade async Python services, LLM/RAG pipelines, and the voice layer where relevant (STT/TTS, telephony). Hybrid work with Madrid/Barcelona offices offered.
Location: Madrid (preferred) or Barcelona (1-2 office days)
Company: autonomous AI agents for enterprise clients
This rapidly scaling start-up is building an AI-native platform that lets enterprise clients run high-volume operations through autonomous, conversational agents — voice, chat and messaging, deployed across dozens of countries in over 60 languages, for some of the largest employers in the world.
The company's core technical challenge is multi-agent orchestration that stays rock-solid at scale — reliable, secure and observable whether it's deployed for a healthcare provider in the US or a retailer in Latin America.
Founded by an experienced team with two prior successful exits, backed by a top-tier European VC, and already trusted by household-name enterprise customers.
You'll build, operate and continuously improve the reliability of the agentic AI systems running in production for enterprise clients — focused on how agents behave, decide, fail, recover and scale under real operational load, not just on shipping new features.
You'll own the observability and evaluation layer that keeps autonomous agents trustworthy at scale — tracing, logging, cost tracking across model calls, regression testing through prompt and model changes, and alerting on failures, hallucinations and degraded performance before customers ever see them.
You'll also work end-to-end across the stack: production-grade async Python services, LLM/RAG pipelines, and the conversational voice layer where relevant — STT/TTS, telephony, real-time latency.