Senior LLMOps Engineer: Scale GenAI Pipelines & RAG

Steampunk

Bloomington (IL)

On-site

USD 145,000 - 185,000

Full time

5 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Steampunk is seeking an experienced Senior LLMOps Engineer to design, implement, and maintain production-grade LLM and RAG pipelines across enterprise environments. You will build hosting, inference, retrieval, and context management, with a focus on security, reliability, and governance.

The role requires expert Python, hands-on experience with Hugging Face Transformers, vLLM, and vector databases, plus CI/CD for AI workloads.

Qualifications

  • U.S. government security clearance capability is required or achievable.
  • Bachelor’s degree with 10+ years total experience.
  • 5+ years in software/data engineering, MLOps, or cloud engineering with 2+ years on LLM/GenAI ops.
  • Hands-on RAG pipeline design, build and maintenance experience.
  • Fluency in Python and production-grade model deployment experience.

Responsibilities

  • Architect and maintain scalable LLM and RAG pipelines with hosting, inference, and context management.
  • Design secure GenAI infrastructure across cloud environments for reliability and cost efficiency.
  • Build automated evaluation systems for quality, safety, latency and governance compliance.
  • Develop CI/CD workflows for AI applications including dataset versioning and model lineage.
  • Collaborate with AI Product Engineers to productionize prototypes into enterprise-grade systems.
  • Integrate vector databases, model gateways, content filters, and guardrails end-to-end.
  • Implement observability to track performance, hallucinations, costs, and usage patterns.
  • Lead troubleshooting and RCA for deployment and inference issues.
  • Produce Python-based services and integrations to support LLM/RAG pipelines.
  • Stay current with MLSecOps trends and governance practices.

Skills

Security clearance
Software engineering
MLOps
LLM/GenAI operations
RAG pipelines
Python
CI/CD for AI/ML
Observability & monitoring
Technical leadership
Cloud architectures
Docker & Kubernetes
Terraform/CloudFormation
AI governance & safety
Agile project mgmt

Education

Bachelor’s degree

Tools

Hugging Face Transformers
vLLM
TensorRT-LLM
LangChain
LlamaIndex
FAISS
Milvus
Pinecone
FastAPI
PyTorch

Job description

Steampunk is seeking an experienced Senior LLMOps Engineer to design, implement, and maintain production-grade LLM and RAG pipelines across enterprise environments. You will build hosting, inference, retrieval, and context management, with a focus on security, reliability, and governance.

The role requires expert Python, hands-on experience with Hugging Face Transformers, vLLM, and vector databases, plus CI/CD for AI workloads.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior LLMOps Engineer: GenAI Infra & RAG Architect
Senior LLMOps Engineer: GenAI Infra & RAG Architect

Steampunk, Inc. • McLean (VA)

On-site
USD 145,000 - 185,000
LLMOps Engineer: Scale, Secure & Observe GenAI Pipelines
LLMOps Engineer: Scale, Secure & Observe GenAI Pipelines

UNAVAILABLE • McLean (VA)

On-site
USD 150,000 - 210,000
Senior LLMOps Engineer - Scalable GenAI Platform Lead
Senior LLMOps Engineer - Scalable GenAI Platform Lead

UNAVAILABLE • McLean (VA)

On-site
USD 180,000 - 240,000
Senior LLMOps Engineer
Senior LLMOps Engineer

UNAVAILABLE • McLean (VA)

On-site
USD 180,000 - 240,000
Senior AI/ML Engineer: Build Production-Ready GenAI Apps
Senior AI/ML Engineer: Build Production-Ready GenAI Apps

Steampunk • McLean (VA)

Hybrid
USD 140,000 - 190,000
GenAI Production Engineer: RAG, LLMs & ML Ops
GenAI Production Engineer: RAG, LLMs & ML Ops

HireHi • United States

Remote
USD 120,000 - 180,000
Senior AI Lead: Enterprise ML & MLOps
Senior AI Lead: Enterprise ML & MLOps

Steampunk • McLean (VA)

Hybrid
USD 140,000 - 180,000
Senior AI/ML Full-Stack Engineer: RAG & LLM Orchestration
Senior AI/ML Full-Stack Engineer: RAG & LLM Orchestration

MAS Global Consulting • Plano (TX)

On-site
USD 140,000 - 190,000
Senior AI Engineer: Scalable AI & MLOps Lead
Senior AI Engineer: Scalable AI & MLOps Lead

Talentify • Austin (TX)

On-site
USD 140,000 - 220,000
Senior AI Engineer: Scalable AI & MLOps Lead
Senior AI Engineer: Scalable AI & MLOps Lead

Talentify • Austin (TX)

On-site
USD 140,000 - 220,000