Senior AI Engineer

Gateworth Group

Dubai

On-site

AED 420,000 - 660,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Family benefits

Job summary

Gateworth Group is expanding its AI engineering capability in the UAE and building a team focused on large‑scale model performance. Senior engineers will work deeply in the internals of modern LLMs, optimise inference behaviour, and contribute to high‑performance model deployment across complex environments.

This role suits someone who enjoys highly technical work and understands transformer‑based models at scale, moving across research, engineering, and system‑level optimisation to shape

Responsibilities

  • Analyse, profile, and optimise LLM inference performance across distributed, multi‑chip or multi‑node systems
  • Apply deep understanding of transformer architectures, including dense and MoE models
  • Evaluate and benchmark leading LLMs (LLaMA, Mistral, Qwen, DeepSeek) across different hardware environments
  • Design and implement optimisations for attention mechanisms such as Flash Attention, grouped‑query attention, and sliding‑window attention
  • Work on model‑level optimisation techniques including quantisation (INT8/FP8), KV‑cache management, batching, and parallelism strategies
  • Collaborate with hardware, systems, and compiler teams to co‑design efficient inference pipelines
  • Build and maintain benchmarking frameworks to evaluate latency, throughput, and scaling behaviour
  • Analyse trade‑offs between model architecture choices and system‑level performance, contributing to deployment strategies for large‑scale environments
  • Stay current with research in LLM architectures and inference optimisation

Skills

Transformer architectures
LLMs (LLaMA, Mistral, Qwen, DeepSeek)
MoE architectures
Attention mechanisms optimization
Inference optimization
Python & PyTorch/JAX
Distributed systems
Profiling performance
Systems thinking

Job description

Position: Senior AI Engineer – LLM Systems

Compensation: Competitive salary plus family benefits & variable

Location: Dubai

Overview

Gateworth Group is supporting a technology organisation in the UAE that is expanding its AI engineering capability and building out a specialised team focused on large‑scale model performance. The company is investing heavily in advanced AI systems and is hiring senior engineers who can work deep in the internals of modern LLMs, optimise inference behaviour, and contribute to high‑performance model deployment across complex environments.

This role suits someone who enjoys highly technical work, understands how transformer‑based models behave at scale, and can move comfortably between research, engineering, and system‑level optimisation. You’ll work across model analysis, performance tuning, benchmarking, and architecture decisions, helping shape next‑generation AI infrastructure.

Main Responsibilities
  • Analyse, profile, and optimise LLM inference performance across distributed, multi‑chip or multi‑node systems
  • Apply deep understanding of transformer architectures, including dense and Mixture‑of‑Experts (MoE) models
  • Evaluate and benchmark leading LLMs (LLaMA, Mistral, Qwen, DeepSeek) across different hardware environments
  • Design and implement optimisations for attention mechanisms such as Flash Attention, grouped‑query attention, and sliding‑window attention
  • Work on model‑level optimisation techniques including quantisation (INT8/FP8), KV‑cache management, batching, and parallelism strategies
  • Collaborate with hardware, systems, and compiler teams to co‑design efficient inference pipelines
  • Build and maintain benchmarking frameworks to evaluate latency, throughput, and scaling behaviour
  • Analyse trade‑offs between model architecture choices and system‑level performance, contributing to deployment strategies for large‑scale environments
  • Stay current with research in LLM architectures and inference optimisation
Qualifications
  • Strong understanding of transformer architectures and LLM internals
  • Hands‑on experience with multiple modern LLMs (LLaMA, Mistral, Qwen, DeepSeek)
  • Deep knowledge of dense and Mixture‑of‑Experts (MoE) architectures
  • Familiarity with attention mechanisms and their optimisation strategies
  • Experience with inference optimisation techniques (quantisation, pruning, KV‑caching, batching)
  • Strong Python skills and experience with ML frameworks such as PyTorch or JAX
  • Experience with distributed systems and large‑scale inference workloads
  • Ability to profile and debug performance bottlenecks across hardware and software stacks
  • Strong systems thinking with the ability to work across model, runtime, and hardware layers
Preferred:
  • 8+ years’ experience in deep learning, AI systems, or performance engineering
  • Experience working close to hardware
  • Familiarity with parallelism strategies (tensor, pipeline, expert parallelism)
  • Experience with datacenter‑scale deployment and inference servers (e.g., vLLM)
  • Background in performance engineering or systems optimisation
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI LLM Engineer
AI LLM Engineer

DiceTek UAE • Dubai

On-site
Exposure to advanced AI technologies
Opportunity for career growth in AI engineering
Work in a dynamic tech industry
Senior LLM Performance Engineer - Inference (Dubai)
Senior LLM Performance Engineer - Inference (Dubai)

Gateworth Group • Dubai

On-site
AED 420,000 - 660,000
Family benefits
Principal AI Engineer
Principal AI Engineer

Anson McCade • Dubai

Hybrid
AED 300,000 - 520,000
Remote work across UAE
Learning & development budget
Home office equipment allowance
+2
Senior AI Engineer
Senior AI Engineer

Technology Innovation Institute • United Arab Emirates

On-site
AED 300,000 - 450,000
senior AI Engineer
senior AI Engineer

Technology Innovation Institute • Dubai

On-site
AED 400,000 - 650,000
Senior Machine Learning Engineer (m/f/d)
Senior Machine Learning Engineer (m/f/d)

Halian • Abu Dhabi

On-site
AED 300,000 - 450,000
Associate AI software Engineer
Associate AI software Engineer

NEXUS BLUE PROJECT MANAGEMENT • Dubai

On-site
Travel allowance
Visa sponsorship
Insurance
ML / LLMOps Engineer
ML / LLMOps Engineer

SentraAI • Dubai

On-site
AED 250,000 - 350,000
AI Engineer (LLM, RAG & Intelligent Automation) | Shaz MLC | Dubai, UAE
AI Engineer (LLM, RAG & Intelligent Automation) | Shaz MLC | Dubai, UAE

Shaz MLC • Dubai

On-site
AED 180,000 - 240,000
Senior AI Engineer
Senior AI Engineer

D4 Insight • Abu Dhabi

On-site
AED 350,000 - 550,000