Remote AI Systems Engineer: Kernel & Inference Optimization

Jobgether SRL

Abu Dhabi

Remote

AED 250,000 - 360,000

Full time

10 days ago
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Remote-first environment
International team
Exposure to cutting-edge AI research

Job summary

Jobgether SRL is seeking an AI Research Engineer (Kernel & Inference Optimization) in the United Arab Emirates to advance model serving for edge and mobile platforms. You will bridge AI research with systems engineering to optimize latency, throughput, and memory usage across diverse hardware.

Responsibilities include designing high-throughput inference pipelines, benchmarking, and developing custom GPU kernels. A PhD in NLP/ML or CS degree with strong research output is highly valued.

Qualifications

  • Degree in Computer Science or a related technical field; a PhD in NLP, Machine Learning, or a related discipline is highly relevant, particularly with a strong AI research track record and publications at leading conferences.
  • Proven expertise in Metal Shading Language (MSL), including the ability to write custom compute shaders from scratch.
  • Demonstrated experience with low-level kernel optimization and inference optimization on mobile or other resource-constrained devices.
  • Track record of delivering measurable improvements in inference latency, throughput, and memory footprint for domain-specific applications.
  • Deep understanding of modern model-serving architectures, inference engines, and optimization techniques for high-performance AI deployment.
  • Strong experience writing GPU kernels for mobile devices.

Responsibilities

  • Design and deploy advanced model-serving architectures optimized for high throughput, low latency, and efficient memory utilization.
  • Develop inference pipelines capable of operating effectively across diverse environments, including resource-constrained mobile devices and edge platforms.
  • Establish clear performance targets covering response latency, token generation speed, throughput, memory footprint, and reliability.
  • Build and execute controlled inference benchmarks in simulated and production environments, tracking latency, throughput, memory consumption, and error rates.
  • Create and maintain representative datasets and simulation scenarios for evaluating model performance under real-world and resource-constrained conditions.
  • Identify computational and memory bottlenecks across inference pipelines and implement solutions involving batching, networking, memory management, and other system-level optimizations.
  • Develop custom GPU kernels and compute shaders for mobile hardware, including solutions written in Metal Shading Language (MSL).
  • Apply advanced inference optimization techniques such as pruning, quantization, Flash Attention, KV caching, and speculative decoding.
  • Design and optimize distributed inference systems using approaches such as tensor parallelism, pipeline parallelism, and expert parallelism for large-scale GPU workloads.
  • Work with cross-functional engineering and research teams to integrate optimized inference frameworks into production and edge-device applications.
  • Define evaluation methodologies, document experimental results, compare performance against established benchmarks, and continuously refine optimization strategies.
  • Monitor production performance and use empirical research to identify opportunities for further improvements in scalability, efficiency, and reliability.

Skills

MSL
Low-level kernel optimization
Inference optimization
GPU kernels for mobile
Model serving architectures
Diffusion models
Vision Transformers
Empirical benchmarking
Distributed inference
English communication

Education

PhD in NLP/ML
Degree in Computer Science

Tools

MSL
Mobile GPU programming

Job description

Jobgether SRL is seeking an AI Research Engineer (Kernel & Inference Optimization) in the United Arab Emirates to advance model serving for edge and mobile platforms. You will bridge AI research with systems engineering to optimize latency, throughput, and memory usage across diverse hardware.

Responsibilities include designing high-throughput inference pipelines, benchmarking, and developing custom GPU kernels. A PhD in NLP/ML or CS degree with strong research output is highly valued.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Inference Architect: Kernel & Edge Optimizations
AI Inference Architect: Kernel & Edge Optimizations

Jobgether SRL • United Arab Emirates

Remote
AED 350,000 - 700,000
Remote-first
International team
Cutting-edge research
+2
AI Research Engineer (Kernel & Inference Optimization)
AI Research Engineer (Kernel & Inference Optimization)

Jobgether SRL • United Arab Emirates

Remote
AED 350,000 - 700,000
Remote-first
International team
Cutting-edge research
+2
AI Platform & LLM Infra Engineer
AI Platform & LLM Infra Engineer

iSystems Ltd • Dubai

On-site
AED 279,000 - 502,000
AI Research Engineer (Kernel & Inference Optimization)
AI Research Engineer (Kernel & Inference Optimization)

Jobgether SRL • Abu Dhabi

Remote
AED 250,000 - 360,000
Remote-first environment
International team
Exposure to cutting-edge AI research
Remote AI-Driven Full-Stack Engineer
Remote AI-Driven Full-Stack Engineer

Hired • United Arab Emirates

On-site
AED 150,000 - 230,000
Senior AI Scientist: Deep Learning & GPU-Optimized Production
Senior AI Scientist: Deep Learning & GPU-Optimized Production

Avrioc Technologies • Abu Dhabi

On-site
AED 240,000 - 320,000
AI Engineer: Architect & Deploy Multimodal AI Systems
AI Engineer: Architect & Deploy Multimodal AI Systems

AGAPI Technologies • United Arab Emirates

On-site
AED 180,000 - 280,000
Competitive compensation
Visa processing and UAE benefits
Generous holidays
+3
Dubai-based Senior AI Engineer—LLM Inference & Performance
Dubai-based Senior AI Engineer—LLM Inference & Performance

HireHouse • Dubai

On-site
AED 400,000 - 800,000
Senior AI Engineer Gateworth Group On-site Fast Track available
Senior AI Engineer Gateworth Group On-site Fast Track available

HireHouse • Dubai

On-site
AED 400,000 - 800,000
Senior AI-Native Product Engineer (Remote)
Senior AI-Native Product Engineer (Remote)

Jobgether SRL • Abu Dhabi

Remote
AED 350,000 - 550,000
Fully remote
Competitive compensation
Vacation 28 days
+9