Responsibilities
- Design, build and maintain AI pipelines and automations end‑to‑end for data, model serving, and monitoring.
- Integrate and orchestrate LLMs and generative models via APIs (Anthropic, OpenAI, DeepSeek, and others) as well as open‑source models.
- Fine‑tune / adapt models (LoRA and full fine‑tuning); perform dataset preparation, training, evaluation, and serving.
- Deploy and operate services in production on servers, GPUs, containers, CI/CD pipelines, with monitoring and safe rollbacks.
- Build backend logic (primarily Python), basic frontend components, and work with databases.
- Leverage AI coding agents (Claude Code, OpenAI Codex, Cursor, Cline, and similar) to ship features quickly.
- Continuously research and adopt the newest AI tools, models, and techniques.
Requirements (must‑have)
- Fine‑tuning experience – hands‑on LoRA or full fine‑tuning of LLMs or vision/diffusion models (data → train → evaluate → serve). Mandatory.
- Deployment / MLOps experience – shipping to production on Linux/GPU servers (Docker, systemd/CI, git), with monitoring and rollback discipline. Mandatory.
- Strong programming skills (Python primary) with backend fundamentals, and basic frontend (HTML/JS, React or similar).
- Multi‑provider LLM API integration – Anthropic, OpenAI, DeepSeek, etc. (API keys, structured JSON outputs, token/cost management, retries/error handling).
- Open‑source model knowledge – running/serving models locally (e.g., LLaMA, Mistral, Qwen, Stable Diffusion/FLUX), quantization/inference basics.
- Databases / SQL (Postgres or similar).
- Proficiency with AI coding agents – Claude Code, Codex, Cursor, Cline, and other AI dev tools – used effectively for real engineering work.
- Strong research mentality – stays current with the daily, fast‑moving AI landscape and can independently evaluate and adopt new tools/models.
- Strong problem‑solving, attention to detail, and production discipline.
Nice‑to‑have (bonus)
- ComfyUI / diffusion workflows, RAG & vector databases, agent / tool‑use frameworks, prompt‑engineering depth, media/ffmpeg pipelines, cloud/GPU infrastructure, distributed‑systems patterns (queues, idempotency, checkpointing).
Working style (it matters for live production)
- Ownership & autonomy – takes a task, researches, plans, executes, and documents it end‑to‑end.
- Production rigor – verifies before calling anything "done," always backs up + has a rollback.
- Clear communicator – explains technical decisions in plain English to non‑engineers.
- Fast but careful – moves quickly without cutting corners on safety.
Location & Details
Business Bay, Dubai, UAE.
Job Type: Full‑time, On‑site (no remote option).
Working Hours: 12:00 PM – 9:00 PM, Monday to Saturday. Aligns with UK, Europe, and Americas time zones.
Salary: AED 4,000 – 6,000 per month + AED 350 travel allowance + Visa + Insurance.
Experience Required: 3–5 years.