Staff AI Engineer

Wati Dot I O

Hong Kong

On-site

HKD 900,000 - 1,500,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Direct access to founding team
Ownership of AI roadmap

Job summary

Wati is hiring a Staff AI Engineer to own LLM orchestration, RAG, and agent infrastructure at a 4B+ messages/year scale. You will lead architecture, deployment, and optimization of LLM-driven services, including multi-provider inference, RAG pipelines, multi-agent workflows, and voice AI, in a senior IC role with significant technical influence.

We seek a builder who bridges complex AI capabilities with massive production environments to ensure fast, reliable, and cost-effective AI across

Qualifications

  • 5+ years in backend or infrastructure engineering with strong AI exposure.
  • Expertise in at least one high-performance language and Python proficiency.
  • Experience deploying LLM/NLP models to production at scale.

Responsibilities

  • Architect and lead AI production stack across providers.
  • Design scalable RAG pipelines and tool-calling infrastructure.
  • Guide voice and multimodal AI integration across channels.
  • Own data collection, cleaning, and model versioning pipelines.
  • Optimize API costs, latency, and caching for large-scale workloads.
  • Build AI quality assessment infrastructure with production metrics.
  • Drive technology decisions and architectural patterns with leadership.

Skills

Go
Rust
C++
Python
LLM deployment
Data pipelines
Vector DBs
Kubernetes
Infra as Code
Multi-provider orchestration

Tools

GCP
AWS
Docker
LiveKit
WebRTC

Job description

About Wati

Started as a WhatsApp team inbox in 2020, Wati has evolved into an AI-powered customer engagement platform that goes beyond a single channel. Designed for businesses that sell, support, and grow through conversations, Wati observes customer intent in real time, decides the next best revenue action, and executes it across marketing, sales, and support — on WhatsApp, Instagram, Facebook, TikTok, SMS, and more.


Trusted by over 16,000 customers across 190+ countries, Wati simplifies complex operations and business conversations with a unified inbox, no-code automation, and our intelligent AI layer, Astra.


Proudly backed by Tiger Global, Sequoia Capital, DST Global, and Shopify, and recognised as a Premium Partner of Meta and Google.


About the Role

We’re hiring a Staff AI Engineer to own LLM orchestration, RAG, and agent infrastructure at 4B+ messages/year scale.


Our platform processes over 4 billion messages per year across 100+ countries. Your mission is to build the robust, scalable, and intelligent systems that turn conversation data into real-time, intelligent customer experiences.


In this role, you will lead the architecture, deployment, and optimization of our LLM-driven services — including multi-provider inference orchestration, RAG pipelines, multi-agent workflows, and voice AI. This is a senior IC role with significant technical influence across the AI stack.


We need a "builder" who can bridge the gap between complex AI capabilities and massive-scale production environments, ensuring our AI is fast, reliable, and cost-effective.


What You Will Own


  • Core LLM Infrastructure: Architect and lead our AI production stack, including multi-provider LLM gateway optimization, token budget management, and low-latency inference routing across OpenAI, Gemini, and other providers.

  • Agentic AI & RAG: Design and implement scalable RAG (Retrieval-Augmented Generation) systems, multi-step AI agent workflows, and tool-calling infrastructure (MCP), ensuring high accuracy and reliability in customer interactions.

  • Voice & Multimodal AI: Lead the evolution of our voice AI layer (WebRTC/realtime) and cross-channel agent coordination across text, voice, and connected messaging platforms.

  • AI Production Lifecycles: Own the "Engineering-to-AI" loop: building automated pipelines for data collection, cleaning, fine-tuning orchestration, and model versioning.

  • Performance & Cost Optimization: Continuously optimize API costs, token budgets, latency, and caching strategies to ensure our 4-billion-message scale remains sustainable and performant.

  • Evaluation & Benchmarking: Build the infrastructure for systematic AI quality assessment, identifying failure modes and ensuring model improvements are grounded in real-world production metrics.

  • Technical Roadmap: Drive technology decisions in close collaboration with engineering leadership, selecting frameworks and architectural patterns that will define our AI future.


What We Are Looking For


  • Systems Expert: 5+ years of professional experience in backend or infrastructure engineering. Mastery of at least one high-performance language (Go, Rust, or C++) and deep proficiency in Python.

  • AI Deployment Mastery: Proven track record of taking LLMs/NLP models from experiments to high-traffic production. You understand multi-provider orchestration, prompt engineering at scale, and model drift management.

  • Data Pipeline Experience: Strong experience building data pipelines for AI workloads, including document processing, embedding generation, and vector search.

  • Product-Minded Engineer: You don’t just build for the sake of tech; you understand how AI performance impacts customer outcomes and business value.

  • Autonomous Builder: You thrive in environments with high ambiguity and can design, code, and deploy complex systems independently.

  • Experience with vector databases (e.g., Qdrant, Milvus, Pinecone) and RAG architecture patterns.

  • Familiarity with agentic frameworks, tool-calling protocols (MCP, function calling), or multi-agent orchestration.

  • Experience with real-time voice/audio AI pipelines (WebRTC, LiveKit, or similar).

  • Infrastructure-as-Code experience with GCP/AWS, Docker, and Kubernetes.


Benefits

You’ll own AI quality across a platform that serves 16,000+ businesses in 190+ countries. The data pipeline and production infrastructure are in place — your job is to push the frontier: better models, smarter agents, faster inference, and measurable business impact.


You’ll have direct access to the founding team and the autonomy to shape our AI roadmap. This is a rare IC opportunity to own AI end-to-end at production scale, with real data, real customer impact, and a direct line to product decisions.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Staff Applied AI Scientist
Staff Applied AI Scientist

Wati Dot I O • Hong Kong

On-site
HKD 900,000 - 1,300,000
Agentic Engineer - Voice AI
Agentic Engineer - Voice AI

Wati Dot I O • Hong Kong

On-site
HKD 900,000 - 1,300,000
Senior Agentic Engineer - Enterprise
Senior Agentic Engineer - Enterprise

Wati Dot I O • Hong Kong

On-site
HKD 900,000 - 1,300,000
International collaboration
AI-assisted tooling
Direct impact
Agentic Engineer - Backend Software - Shenzhen
Agentic Engineer - Backend Software - Shenzhen

Wati • Hong Kong

On-site
HKD 480,000 - 800,000
Agentic Engineer II - AI Agents (Customer Agents)
Agentic Engineer II - AI Agents (Customer Agents)

Wati Dot I O • Hong Kong

On-site
HKD 600,000 - 900,000
Agentic Engineer II - Voice AI
Agentic Engineer II - Voice AI

Wati Dot I O • Hong Kong

On-site
HKD 942,000 - 1,413,000
Agentic Engineer II - Voice AI
Agentic Engineer II - Voice AI

Wati • Hong Kong

On-site
HKD 700,000 - 1,200,000
Agentic Engineer II - AI Agents (Customer Agents)
Agentic Engineer II - AI Agents (Customer Agents)

Wati • Hong Kong

On-site
HKD 680,000 - 900,000
Staff AI Engineer: Scale LLMs, Orchestrate Agents & RAG
Staff AI Engineer: Scale LLMs, Orchestrate Agents & RAG

Wati Dot I O • Hong Kong

On-site
HKD 900,000 - 1,500,000
Direct access to founding team
Ownership of AI roadmap
Chief Technology Officer | World-class AI product
Chief Technology Officer | World-class AI product

Osmium Consulting Group Limited • Hong Kong

On-site
HKD 1,800,000 - 3,500,000
Significant equity stake
Flexible work-from-home
Work From Anywhere
+1