Senior LLMOps Engineer: Production AI Pipelines

Ncs

Singapore

On-site

SGD 120,000 - 180,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

NCS invites experienced LLMOps Engineers to join NCS AI Central in Singapore, delivering production-grade AI deployments across fast-paced POC/POV to steady-state operations. You will own deployment pipelines, observability, and cost optimization for LLM services, collaborating with AI architects and cloud teams to scale reliable AI systems.

The role spans from POC/POV deployments to production maintenance, with emphasis on reusable patterns, SLA-driven operations, and mentorship of junior

Qualifications

  • 2+ years in DevOps/MLOps/platform engineering, with hands-on exposure to LLM or ML systems in production; 5+ years with end-to-end ownership expected at Senior level.
  • Hands-on with containers, Kubernetes, CI/CD (GitHub Actions/GitLab/Jenkins), and Infrastructure-as-Code (Terraform).
  • Practical experience deploying and operating LLM applications (RAG/agentic systems) at scale, including model gateway/routing patterns.
  • Strong programming ability (Python and/or Go), comfortable across AWS/Azure/GCP; GCC/HCC exposure a plus.
  • Working knowledge of observability tooling (OpenTelemetry, Prometheus/Grafana, ELK/OpenSearch) applied to AI signals.
  • Comfortable in fast-paced POC/POV settings and disciplined, SLA-driven production support.
  • Working knowledge of China AI stack (DeepSeek, Qwen, GLM, Kimi, MiniMax) deployment patterns, licensing, self-hosting.

Responsibilities

  • Model & Deployment Operations
  • Own the deployment pipeline for LLM and agentic applications - versioning models, prompts, and configurations as code, with safe rollout and rollback paths.
  • Build repeatable CI/CD pipelines for AI services (containerised, on Kubernetes) so new model or prompt versions ship without manual intervention.
  • Support rapid, disposable environment spin-up for FDE POC/POV work, then harden the same pipeline into a production-grade deployment when an engagement scales.
  • Monitoring & Reliability
  • Instrument LLM applications with observability for latency, error rates, output drift, and hallucination signals - not just infrastructure uptime.
  • Define and track SLAs/SLOs for production AI services in partnership with AI Architects and Solution Architects.
  • Set up alerting and runbooks so issues in live AI systems are caught and triaged before clients notice.
  • Cost & Performance Management
  • Monitor token spend and inference cost per model/engagement; flag anomalies and right-sizing opportunities with AI FinOps.
  • Tune routing between model tiers for cost-performance balance at production scale.
  • FDE & Development/Maintenance Coverage
  • During FDE engagements: stand up lightweight, reusable deployment scaffolding for rapid iteration without operational overhead.
  • During system development & maintenance: own steady-state production operations, patching, upgrades, and incident response.
  • Contribute reusable deployment patterns back into the internal asset library.
  • Collaboration
  • Work closely with AI Engineers, Cloud Architects, and AI Site-facing leads for smooth handoff from POC to Scale to Operate.
  • Mentor engineers on LLMOps and participate in PRR gate reviews.

Skills

DevOps/MLOps experience
LLM/ML production
End-to-end ownership
Containers
Kubernetes
CI/CD
Terraform
Model routing
Python
Go
Cloud platforms
OpenTelemetry
Prometheus
Grafana
ELK/OpenSearch
Postgres
Redis
pgvector
Pinecone
Weaviate
GitHub Actions
GitLab
Jenkins
vLLM/TGI
Python
Go
Docker
Helm
Argo CD
Terraform
Vault

Tools

Docker
Kubernetes
Helm
Argo CD
Terraform
Vault
OpenTelemetry
Prometheus
Grafana
ELK/OpenSearch
Postgres
Redis
pgvector
Pinecone
Weaviate
GitHub Actions
GitLab
Jenkins
vLLM/TGI
Python
Go

Job description

NCS invites experienced LLMOps Engineers to join NCS AI Central in Singapore, delivering production-grade AI deployments across fast-paced POC/POV to steady-state operations. You will own deployment pipelines, observability, and cost optimization for LLM services, collaborating with AI architects and cloud teams to scale reliable AI systems.

The role spans from POC/POV deployments to production maintenance, with emphasis on reusable patterns, SLA-driven operations, and mentorship of junior

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

LLMOps Architect — Production-Ready AI Systems
LLMOps Architect — Production-Ready AI Systems

NCS • Singapore

On-site
SGD 120,000 - 180,000
Career development
Learning opportunities
Team NCS culture
Senior LLMOps Engineer: Production AI Reliability Lead
Senior LLMOps Engineer: Production AI Reliability Lead

NCS Pte Ltd • Singapore

On-site
SGD 120,000 - 180,000
Health insurance
Learning & development
#EG Senior / LLMOps Engineer
#EG Senior / LLMOps Engineer

NCS • Singapore

On-site
SGD 120,000 - 180,000
Career development
Learning opportunities
Team NCS culture
Senior MLOps Engineer: Build Production-Scale ML Pipelines
Senior MLOps Engineer: Build Production-Scale ML Pipelines

EPAM Systems • Singapore

On-site
SGD 120,000 - 180,000
#EG Senior / LLMOps Engineer
#EG Senior / LLMOps Engineer

NCS Pte Ltd • Singapore

On-site
SGD 120,000 - 180,000
Health insurance
Learning & development
Gen AI Engineer — Fast, Impactful POC to Production
Gen AI Engineer — Fast, Impactful POC to Production

NCS Pte Ltd • Singapore

On-site
SGD 90,000 - 130,000
Senior Data Engineer: AI-Ready Pipelines & RAG
Senior Data Engineer: AI-Ready Pipelines & RAG

Ncs • Singapore

On-site
SGD 120,000 - 180,000
AI Engineer: MLOps, Production Pipelines
AI Engineer: MLOps, Production Pipelines

Seatrium • Singapore

On-site
SGD 60,000 - 100,000
Production AI/LLM Engineer: Prompt Craft & Model Integration
Production AI/LLM Engineer: Prompt Craft & Model Integration

NCS Pte Ltd • Singapore

On-site
SGD 120,000 - 190,000
AI Engineer – Production ML & MLOps Specialist
AI Engineer – Production ML & MLOps Specialist

SEATRIUM (SG) PTE. LTD. • Singapore

On-site
SGD 90,000 - 120,000
Island wide transport provided