Forward Deployed AI Engineer (On-Premise GenAI & Integration) | petrus.sa | Riyadh, Saudi Arabia

petrus.sa

Riyadh

On-site

SAR 260,000 - 420,000

Full time

5 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

On-site role in Riyadh

Job summary

petrus.sa is seeking an exceptional Forward Deployed AI Engineer in Riyadh to install, optimize, and integrate on-prem GenAI platforms within client banking environments. You will work hands-on with OpenShift, Kubernetes, and NVIDIA NIM, delivering secure, GPU-accelerated AI deployments in air-gapped networks.

The role emphasizes GitOps pipelines, strict SLA-driven model inference, and collaboration with client IT and security teams to maintain compliance and resilience in enterprise AI delivery.

Qualifications

  • 4+ years software engineering or DevOps experience.
  • 2+ years in MLOps/LLMOps deploying containerized apps in air-gapped environments.
  • Hands-on Red Hat OpenShift or Kubernetes admin experience.
  • Experience with NVIDIA NIM, DataRobot, Graph DBs is preferred.
  • GitOps with ArgoCD/OpenShift Pipelines; Helm/Kustomize used for configs.

Responsibilities

  • Install, configure, and optimize enterprise AI systems on-premises in client networks.
  • Deploy NVIDIA NIM, LLMs, DataRobot, and Graph Databases on OpenShift/Kubernetes clusters.
  • Package and deploy large container images and model weights into isolated environments.
  • Tune OpenShift pods, configure MIG, and optimize local model inference for SLAs.
  • Build automated GitOps pipelines (ArgoCD/OpenShift Pipelines) for deployment/rollback.
  • Interface with client IT/security to remove deployment blockers and ensure compliance.
  • Manage secret storage and enterprise authentication (Vault/OpenShift Secrets).

Skills

MLOps
LLMOps
OpenShift
Kubernetes
NVIDIA NIM
GPU optimization
GitOps
Containerization
Python/Go

Tools

ArgoCD
OpenShift Pipelines
NVIDIA GPU Operator
Kustomize
Helm
Git

Job description

Position Summary:
  • petrus.sa is seeking an exceptional Forward Deployed AI Engineer specializing in on-premise GenAI and integration to join our technology team in Riyadh, Saudi Arabia.
  • This full-time role serves as our hands-on engineering force responsible for installing, optimizing, and integrating sophisticated Generative AI platforms and graph technologies directly into client banking environments.
  • Combining elements of MLOps, LLMOps, and systems engineering, you will work on-site within secure client networks to transform architectural blueprints into production-ready AI deployments.
  • The ideal candidate brings 4+ years of software engineering or DevOps experience, with 2+ years specialized in deploying accelerated containerized applications inside air-gapped environments.
  • Operating in Riyadh, you will partner directly with client IT and security teams to administer Red Hat OpenShift clusters, NVIDIA microservices, and enterprise data platforms.
  • petrus.sa provides a dynamic, high-impact workplace where your technical leadership directly empowers secure, mission-critical financial artificial intelligence initiatives.
  • We welcome driven systems engineers who excel in complex container orchestration, GPU optimization, and secure enterprise software delivery.
Detailed Job Description:
  • As a Forward Deployed AI Engineer at petrus.sa, your core responsibilities encompass leading the practical installation, configuration, and optimization of enterprise AI systems on-premise.
  • You will deploy and manage NVIDIA NIMs, large language models (LLMs), DataRobot, and enterprise Graph Databases on client-owned OpenShift and Kubernetes clusters.
  • Your daily operational duties involve packaging, mirroring, and deploying massive container images, software dependencies, and model weights into strictly isolated, air-gapped server environments.
  • You will fine-tune OpenShift pods, configure GPU time-slicing and Multi-Instance GPU (MIG), and optimize local model inference to meet strict banking SLAs for token latency and throughput.
  • Constructing robust, fully automated on-premise GitOps pipelines using ArgoCD or OpenShift Pipelines forms a crucial part of your daily software delivery lifecycle.
  • You will act as the trusted technical interface for client IT, security, and infrastructure teams to clear accelerated compute deployment blockers and ensure compliance.
  • Managing secure secret storage, enterprise authentication frameworks, and continuous model regression testing will keep your infrastructure secure and resilient.
Key Responsibilities:
  • Lead the practical installation, configuration, and optimization of NVIDIA NIMs, LLMs, DataRobot, and Graph Databases on client OpenShift and Kubernetes clusters.
  • Package, mirror, and deploy massive container images, software dependencies, and model weights into isolated, air-gapped server environments.
  • Fine-tune OpenShift pods, configure GPU time-slicing/MIG, and optimize local model inference using NVIDIA NIM to meet banking SLAs for token latency (TTFT/throughput).
  • Construct robust, fully automated on-premise GitOps pipelines using ArgoCD or OpenShift Pipelines for continuous model deployment and rollback management.
  • Act as the trusted technical interface for client IT, security, and infrastructure teams to clear accelerated compute deployment blockers.
  • Administer and configure applications on Red Hat OpenShift and Kubernetes, including NVIDIA GPU Operator administration.
  • Implement enterprise authentication frameworks (Active Directory, LDAP, Kerberos, OAuth) and secure secret management via HashiCorp Vault or OpenShift Secrets.
  • Execute rigorous pipeline regression testing, performance benchmarking, and node health monitoring.
  • Collaborate with cross-functional data science and security groups to ensure strict regulatory compliance.
  • Maintain up-to-date technical documentation, deployment runbooks, and architecture blueprints.
Required Qualifications & Skills:
  • 4+ years of professional software engineering or DevOps experience.
  • 2+ years of specialized experience in MLOps/LLMOps deploying accelerated containerized applications inside highly regulated or secure air-gapped environments.
  • Direct, hands-on experience administering and configuring applications on Red Hat OpenShift or Kubernetes.
  • Demonstrated exposure to configuring the NVIDIA GPU Operator (CKA/CKAD certifications are a significant plus).
  • Solid practical experience deploying and managing enterprise software platforms, with a strong preference for NVIDIA NIM, DataRobot, and Graph Database maintenance.
  • Advanced technical skills in writing Helm charts, Kubernetes Operators, Kustomize manifests, and managing configurations via GitOps in offline environments.
  • Solid understanding of enterprise authentication frameworks (Active Directory, LDAP, Kerberos, OAuth) and secure secret management (HashiCorp Vault or OpenShift Secrets).
  • Legal eligibility and physical presence capability to work on-site in Riyadh, Saudi Arabia.
Nice-to-Have Skills:
  • Certified Kubernetes Administrator (CKA) or Certified Kubernetes Application Developer (CKAD) certifications.
  • Prior working experience within banking, financial services, or highly regulated government technology sectors.
  • Advanced proficiency in Python or Go scripting for Kubernetes operator customization and automation.
  • Experience with high-performance storage solutions (Ceph, GlusterFS, or NVMe-oF) for large model weight caching.
  • Familiarity with advanced AI agent frameworks and LLM evaluation benchmarks.
Application Information:
  • Salary/Rate: Competitive Market Standard
  • Deadline: Open / Immediate Hiring
  • Notice Period: Immediate to 30 Days Preferred
  • Contract Duration: Full-Time Permanent Position
Recruitment Pro Tip:

Highlight your hands-on experience with Red Hat OpenShift, NVIDIA NIM microservices, and air-gapped GitOps deployments in your CV to immediately capture the hiring manager’s attention.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

On-Prem GenAI Engineer - OpenShift & GPU Expert
On-Prem GenAI Engineer - OpenShift & GPU Expert

petrus.sa • Riyadh

On-site
SAR 260,000 - 420,000
On-site role in Riyadh
Forward Deployed AI Engineer
Forward Deployed AI Engineer

Crayon Data Pvt Ltd • Riyadh

On-site
SAR 300,000 - 520,000
Forward Deployed AI Engineer
Forward Deployed AI Engineer

Zohorecruit • Riyadh

On-site
SAR 240,000 - 360,000
Lead AI Architect
Lead AI Architect

Crayon Data • Riyadh

On-site
SAR 500,000 - 900,000
AI Engineer – Machine Learning, Generative AI & MLOps | Riyadh, Saudi Arabia
AI Engineer – Machine Learning, Generative AI & MLOps | Riyadh, Saudi Arabia

Master Works • Saudi Arabia

On-site
SAR 180,000 - 280,000
Career development opportunities
AI Engineer
AI Engineer

Confidential • Saudi Arabia

On-site
SAR 180,000 - 280,000
Senior Engineer, Saudi Arabia
Senior Engineer, Saudi Arabia

Wonderful Ltd. • Saudi Arabia

On-site
SAR 180,000 - 300,000
Lead AI Architect
Lead AI Architect

Zohorecruit • Riyadh

On-site
SAR 1,350,000 - 1,800,000
Senior Backend Engineer – AI Applications & OpenShift
Senior Backend Engineer – AI Applications & OpenShift

SoftwareONE Deutschland GmbH • Riyadh

Hybrid
SAR 320,000 - 520,000
Global company culture
Mentor support for start
President’s Club recognition
+2
AI Engineer
AI Engineer

Confidential • Riyadh

On-site
SAR 180,000 - 340,000