Senior LLMOps Engineer - Scalable GenAI Platform Lead

UNAVAILABLE

McLean (VA)

On-site

USD 180,000 - 240,000

Full time

5 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

UNAVAILABLE is seeking an experienced Senior LLMOps Engineer to design and maintain production-grade LLM pipelines, deployment architectures, and monitoring across enterprise environments. This role spans model deployment, evaluation, optimization, and RAG pipelines.

You will lead secure GenAI infrastructure, implement CI/CD for AI workloads, integrate vector databases, and ensure governance, privacy, and cost efficiency while mentoring junior engineers.

Qualifications

  • Ability to obtain and maintain a U.S. government security clearance.
  • Bachelor’s degree and 10+ years of total experience.
  • 5+ years in software engineering, data engineering, MLOps, or cloud engineering, with 2+ years in LLM/GenAI operations.
  • Hands-on experience building RAG pipelines.
  • Fluency in Python.
  • Experience deploying models with HF Transformers, vLLM, TensorRT-LLM.
  • Experience with FastAPI, PyTorch, LangChain, LlamaIndex, vector DBs (FAISS/Milvus/Pinecone).
  • Cloud platforms (AWS/Azure/GCP) with hosting, compute, secure networking.
  • CI/CD pipelines, testing for AI workloads.
  • Docker, Kubernetes, Terraform/CloudFormation.
  • MLSocOps, AI governance, model hardening, content safety monitoring.
  • Logging/observability for high-throughput LLMs.
  • Excellent written and verbal communication.
  • Balance long-term platform thinking with hands-on ops.
  • Experience in Agile environments.

Responsibilities

  • Architect, build, and maintain LLM and RAG pipelines including hosting and retrieval layers.
  • Lead secure GenAI infrastructure across cloud environments for reliability and cost efficiency.
  • Build automated evaluation systems for LLM output quality, safety, latency, and governance.
  • Develop CI/CD workflows for LLM/GenAI apps, including data/versioning and model lineage.
  • Collaborate with AI Product Engineers and Data Scientists to productionize prototypes.
  • Integrate vector databases, model gateways, content filters, and guardrails.
  • Implement observability tracking performance, hallucinations, costs, and user patterns.
  • Lead troubleshooting and root-cause analysis for deployment and pipeline issues.
  • Develop Python-based services and integrations supporting LLM/RAG pipelines.
  • Stay current with MLSecOps patterns and AI governance frameworks.

Skills

Python
Communication
Agile

Education

Bachelor's degree

Tools

Hugging Face Transformers
vLLM
TensorRT-LLM
FastAPI
PyTorch
LangChain
LlamaIndex
FAISS
Milvus
Pinecone
Docker
Kubernetes
Terraform
CloudFormation
AWS
Azure
GCP

Job description

UNAVAILABLE is seeking an experienced Senior LLMOps Engineer to design and maintain production-grade LLM pipelines, deployment architectures, and monitoring across enterprise environments. This role spans model deployment, evaluation, optimization, and RAG pipelines.

You will lead secure GenAI infrastructure, implement CI/CD for AI workloads, integrate vector databases, and ensure governance, privacy, and cost efficiency while mentoring junior engineers.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior LLMOps Engineer: GenAI Infra & RAG Architect
Senior LLMOps Engineer: GenAI Infra & RAG Architect

Steampunk, Inc. • McLean (VA)

On-site
USD 145,000 - 185,000
LLMOps Engineer: Scale, Secure & Observe GenAI Pipelines
LLMOps Engineer: Scale, Secure & Observe GenAI Pipelines

UNAVAILABLE • McLean (VA)

On-site
USD 150,000 - 210,000
Senior LLMOps Engineer
Senior LLMOps Engineer

UNAVAILABLE • McLean (VA)

On-site
USD 180,000 - 240,000
Senior LLMOps Engineer: Scale GenAI Pipelines & RAG
Senior LLMOps Engineer: Scale GenAI Pipelines & RAG

Steampunk • Bloomington (IL)

On-site
USD 145,000 - 185,000
Senior LLMOps Engineer - Remote (Production AI)
Senior LLMOps Engineer - Remote (Production AI)

CINC Systems • United States

On-site
USD 150,000 - 210,000
Senior AI Architect & Production LLM Engineer
Senior AI Architect & Production LLM Engineer

UNAVAILABLE • McLean (VA)

On-site
USD 140,000 - 210,000
Senior LLMOps Engineer - Remote (Production AI)
Senior LLMOps Engineer - Remote (Production AI)

CINC Systems • United States

On-site
USD 150,000 - 210,000
Senior AI Engineer: Scalable LLMs & MLOps Leader
Senior AI Engineer: Scalable LLMs & MLOps Leader

Compunnel, Inc. • San Francisco (CA)

On-site
USD 120,000 - 160,000
Senior MLOps Engineer: Scalable AI Infra & Deployment
Senior MLOps Engineer: Scalable AI Infra & Deployment

PathAI • Boston (MA)

Hybrid
USD 128,000 - 196,000
Senior MLOps Engineer – Remote Production AI & LLMs
Senior MLOps Engineer – Remote Production AI & LLMs

Dynatron • United States

Remote
USD 150,000 - 230,000
Remote environment
Professional development
Ownership of production systems
+1