Turn this role into an interview — a resume and cover letter built around what this employer wants.
AethexAI is building a latency‑aware voice platform and seeks an experienced infrastructure engineer to own GPU‑driven model serving at scale. You will optimize fleets for cost per minute while maintaining strict latency budgets across regions.
Based in London and working on an in‑office team of ~10, you will drive reliability, observability, and fast architectural decisions to keep the pipeline responsive and secure.
AethexAI is building a latency‑aware voice platform and seeks an experienced infrastructure engineer to own GPU‑driven model serving at scale. You will optimize fleets for cost per minute while maintaining strict latency budgets across regions.
Based in London and working on an in‑office team of ~10, you will drive reliability, observability, and fast architectural decisions to keep the pipeline responsive and secure.