Senior MLOps Engineer: Build Self-Serve AI Platform

Fathom.io

Dhahran Compound

On-site

SAR 260,000 - 420,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Fathom.io is seeking a mid-to-senior MLOps engineer to build the intelligence layer of our AI platform in Dhahran, Saudi Arabia. You will design infrastructure for model deployment, training, notebooks, UI-driven agents, and RAG pipelines, enabling self-service workflows across GPU and edge deployments.

You’ll work at the intersection of platform engineering, ML infrastructure, distributed systems, and developer experience, collaborating with backend and AI teams to deliver reliable, scalable

Qualifications

  • Strong experience in MLOps, AI platform engineering, machine learning infrastructure, or distributed systems.
  • Hands-on Kubernetes experience, including deploying and operating stateful, training, notebooks, serverless, or GPU-intensive workloads.
  • Experience building ML platforms that support training, experimentation, notebooks, feature/data workflows, model registries, and production serving.
  • Experience with model serving frameworks such as KServe, vLLM, Triton, Ray Serve, or similar.
  • Familiarity with ML lifecycle tooling such as MLflow, experiment tracking, model registries, and CI/CD or GitOps for ML systems.
  • Practical experience optimizing training and inference workloads for latency, throughput, availability, and cost.
  • Understanding of GPU scheduling, resource allocation, model quantization, batching, autoscaling, and multi-model serving.
  • Experience with RAG systems, LLM applications, AI agents, or their supporting infrastructure.
  • Strong software engineering fundamentals and a product-minded approach to platform design.
  • Ability to make complex operational workflows reliable and approachable for users.

Responsibilities

  • Design and build the infrastructure powering our Intelligence layer.
  • Enable reliable, automated workflows for model training, deployment, lifecycle management, and inference.
  • Build scalable foundations for users to create, configure, and operate AI agents and RAG pipelines through the platform UI.
  • Develop the platform capabilities behind managed notebooks, functions, experiments, training jobs, model registries, and serving endpoints.
  • Improve model serving, observability, versioning, evaluation, promotion, and rollback capabilities.
  • Optimize GPU inference and training deployments for performance, reliability, and cost efficiency.
  • Explore efficient approaches for deploying models across centralized GPU infrastructure and edge devices.
  • Automate workflows that otherwise require manual MLOps effort, with a focus on safe, self-service capabilities for platform users.
  • Partner with backend, product, and AI teams to turn complex infrastructure into intuitive platform features.
  • Help define standards for security, multi-tenancy, resource isolation, model governance, and operational reliability.

Skills

MLOps
Distributed systems
Platform engineering
Rust (advantage)
Kubernetes
KServe

Tools

Kubernetes
Knative
MLflow
LangFuse
GPU infrastructure

Job description

Fathom.io is seeking a mid-to-senior MLOps engineer to build the intelligence layer of our AI platform in Dhahran, Saudi Arabia. You will design infrastructure for model deployment, training, notebooks, UI-driven agents, and RAG pipelines, enabling self-service workflows across GPU and edge deployments.

You’ll work at the intersection of platform engineering, ML infrastructure, distributed systems, and developer experience, collaborating with backend and AI teams to deliver reliable, scalable

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

MLOps Platform Engineer: Self-Service AI Infra
MLOps Platform Engineer: Self-Service AI Infra

Fathom.io • Dhahran Compound

On-site
SAR 300,000 - 600,000
MLOps Platform Engineer - Build AI Workflows
MLOps Platform Engineer - Build AI Workflows

Fathom.io • Saudi Arabia

On-site
SAR 240,000 - 480,000
Senior MLOps & LLMOps Engineer
Senior MLOps & LLMOps Engineer

datascience • Riyadh

On-site
SAR 260,000 - 380,000
Senior ML Platform Engineer, Inference & MLOps
Senior ML Platform Engineer, Inference & MLOps

Mollkom • Riyadh

Hybrid
SAR 150,000 - 210,000
Production AI Engineer - On-Site Riyadh (ML & MLOps)
Production AI Engineer - On-Site Riyadh (ML & MLOps)

Master Works • Saudi Arabia

On-site
SAR 180,000 - 280,000
Career development opportunities
Senior ML Platform Engineer - Inference & MLOps (Hybrid)
Senior ML Platform Engineer - Inference & MLOps (Hybrid)

Mollkom • Riyadh

Hybrid
SAR 150,000 - 210,000
MLOps + DevOps Engineer - Agentic AI & Platform
MLOps + DevOps Engineer - Agentic AI & Platform

SYNC • Saudi Arabia

On-site
SAR 280,000 - 560,000
Senior ML Scientist - Financial Forecasting & MLOps
Senior ML Scientist - Financial Forecasting & MLOps

Silver Edge Arabia • Al Khobar

Hybrid
SAR 240,000 - 360,000
Senior Engineer, Saudi Arabia
Senior Engineer, Saudi Arabia

Wonderful Ltd. • Saudi Arabia

On-site
SAR 180,000 - 300,000
AI Engineer – Machine Learning, Generative AI & MLOps | Riyadh, Saudi Arabia
AI Engineer – Machine Learning, Generative AI & MLOps | Riyadh, Saudi Arabia

Master Works • Saudi Arabia

On-site
SAR 180,000 - 280,000
Career development opportunities