AI Systems Architect - LLM & Vector Infrastructure

Starmarkets

Riyadh

On-site

SAR 300,000 - 520,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Starmarkets seeks a senior AI Systems Architect to design AI-native cores for web and mobile apps, leveraging LLMs, vector databases, and agent frameworks. You will shape scalable AI pipelines, RAG systems, memory architectures, and orchestration workflows that power our product stack.

You will drive production-grade AI engineering, ensure security, and build observable, cost-efficient systems with Docker, Kubernetes, and GPU optimization.

Qualifications

  • 5+ years software engineering experience.
  • 2+ years building production AI systems.
  • Deep knowledge of vector embeddings and similarity search.
  • Experience with RAG architectures and tokenization/context window optimization.
  • Experience with Python, FastAPI, and backend services; API design.

Responsibilities

  • Design AI-first architecture for web and mobile apps.
  • Build RAG pipelines using vector databases.
  • Integrate with LLMs (OpenAI, Anthropic, LLaMA) and optimize cost/performance.
  • Develop autonomous AI agents and orchestration flows (n8n, LangGraph/LangChain).
  • Ensure security, monitoring, and guardrails for AI pipelines.

Skills

Software engineering
Production AI systems
RAG architectures
Vector embeddings
Context window optimization
Prompt evaluation frameworks
Distributed systems
Microservices

Tools

Python
FastAPI
PostgreSQL
pgvector
Docker
Kubernetes
GPU optimization
OpenAI APIs
Anthropic
LLaMA

Job description

We are seeking a senior AI Systems Architect to design and implement AI-native application cores where Large Language Models (LLMs), vector databases, retrieval systems, and agent frameworks form the primary computational layer of our web and mobile applications.

This role is responsible for architecting scalable AI pipelines, retrieval-augmented generation (RAG) systems, memory architectures, AI agents, and orchestration workflows integrated with our development stack (Web, Mobile, n8n automation, and AI services).

The ideal candidate understands that AI is not a feature, it is the operating system of the product.

Key Responsibilities
1. AI Core Architecture Design
  • Design AI-first system architecture for web and mobile applications
  • Architect RAG pipelines using vector databases
  • Define long-term memory, short-term memory, and contextual state systems
  • Implement multi-agent AI systems
  • Design AI orchestration layers
2. Vector Database & Embedding Systems
  • Select and implement vector databases such as:
    • Pinecone
    • Weaviate
    • Qdrant
    • Milvus
    • Supabase (pgvector)
  • Optimize embedding strategies
  • Implement hybrid search (semantic + keyword)
  • Design scalable indexing pipelines
3. LLM Integration & Optimization
  • Work with models such as:
    • OpenAI APIs
    • Anthropic
    • Meta (LLaMA)
    • DeepSeek
    • Alibaba (Qwen)
  • Implement structured output pipelines
  • Design evaluation and prompt testing frameworks
  • Optimize cost-performance ratio
4. AI Agent Systems & Orchestration
  • Build autonomous AI agents
  • Design tool-calling systems
  • Integrate with:
    • n8n
    • LangGraph / LangChain style agent flows
  • Implement memory-aware agents
5. Production AI Engineering
  • Build monitoring systems for hallucination detection
  • Design guardrails and validation layers
  • Implement evaluation datasets and benchmarking
  • Ensure security of AI pipelines
  • Build scalable infrastructure (Docker, Kubernetes, GPU optimization)
Technical Expertise
  • 5+ years software engineering experience
  • 2+ years building production AI systems
  • Deep knowledge of:
    • Vector embeddings & similarity search
    • RAG architectures
    • Tokenization and context window optimization
    • Fine-tuning & LoRA concepts
    • Prompt evaluation frameworks
  • Experience with Python (mandatory)
  • Experience with FastAPI / backend services
  • Experience designing scalable APIs
Architecture Experience
  • Designing distributed systems
  • Microservices & event-driven architecture
  • Experience with PostgreSQL + pgvector
  • Experience deploying LLM systems in production
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior AI Systems Architect: LLM, Vectors & Orchestration
Senior AI Systems Architect: LLM, Vectors & Orchestration

Starmarkets • Riyadh

On-site
SAR 300,000 - 520,000
AI Engineer
AI Engineer

Latitude • Riyadh

On-site
SAR 420,000 - 660,000
AI Engineer
AI Engineer

Saudi Azm عزم السعودية • Riyadh

On-site
SAR 260,000 - 460,000
AI Architecture Manager
AI Architecture Manager

Accenture Middle East • Saudi Arabia

On-site
SAR 450,000 - 750,000
AI Engineer - Agentic
AI Engineer - Agentic

Master Works • Riyadh

On-site
SAR 240,000 - 360,000
AI & Machine Learning Consultant
AI & Machine Learning Consultant

Accenture Middle East • Riyadh

On-site
SAR 250,000 - 500,000
AI & Machine Learning Consultant
AI & Machine Learning Consultant

Accenture • Riyadh

On-site
SAR 240,000 - 360,000
AI Architecture Senior Manager
AI Architecture Senior Manager

Accenture Middle East • Riyadh

On-site
SAR 400,000 - 720,000
Principal Forward Deployed Engineer (relocation to Germany)
Principal Forward Deployed Engineer (relocation to Germany)

EPAM Systems • Saudi Arabia

On-site
SAR 360,000 - 600,000
AI Solutions Architect
AI Solutions Architect

JODAYN | جودين • Riyadh

On-site
SAR 300,000 - 400,000