Production AI Engineer: RAG, LLMs & Full-Stack Systems

Oteemo

Reston, Northern (VA, KY)

Hybrid

USD 140,000 - 190,000

Full time

5 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Oteemo, a leading technology consulting firm in Reston, VA, seeks an exceptional AI Software Engineer to build and scale enterprise AI applications end to end, from database to UI. You will work with cutting-edge LLM tech, RAG systems, and production ML infrastructure, blending full-stack development with hands-on AI/ML engineering.

You will design end-to-end RAG pipelines, integrate with GPT-4/Claude/Gemini, and develop prompt strategies, agent systems, and scalable back-end/frontend services.

Qualifications

  • Expert-level proficiency in Python with modern frameworks (FastAPI, Flask).
  • Solid understanding of relational (PostgreSQL, MySQL) and NoSQL (MongoDB) databases.
  • Experience with authentication systems (OAuth2, JWT, SSO) and security best practices.
  • Proven track record of shipping high-quality, scalable software to production.
  • Hands-on experience building and deploying AI/ML applications in production environments.
  • Deep understanding of LLM integration, prompt engineering, and context management.
  • Proven expertise with RAG systems, including document processing, chunking, embedding, retrieval, and generation.
  • Experience working with vector databases (Pinecone, Weaviate, Chroma, FAISS, or Qdrant).
  • Strong grasp of semantic search, similarity algorithms, and hybrid search techniques.
  • Knowledge of evaluation frameworks for assessing AI system quality and performance.
  • Production experience with Docker containerization and Kubernetes orchestration.
  • Strong knowledge of at least one major cloud platform (AWS, Azure, or GCP) and its AI services.
  • Experience building CI/CD pipelines for ML/AI application.
  • Proficiency with infrastructure as code tools (Terraform, CloudFormation, Pulumi).
  • Understanding of monitoring, logging, and alerting best practices; cost optimization experience for cloud and AI workloads.
  • Strong computer science fundamentals and algorithmic thinking.
  • Proficiency with Git workflows, code review practices, and collaborative development.
  • Excellent debugging and problem-solving skills.
  • Clear technical communication and documentation abilities.

Responsibilities

  • Design and implement end-to-end RAG pipelines for intelligent document search and question-answering across enterprise knowledge bases.
  • Build production-ready integrations with leading LLMs (GPT-4, Claude, Gemini) for accurate, contextual responses to user queries.
  • Develop prompt engineering strategies and evaluation frameworks to ensure consistent, high-quality AI outputs.
  • Create agent systems with tool integration capabilities that can autonomously complete complex tasks.
  • Implement vector search solutions using Pinecone, Weaviate, or similar technologies for semantic similarity and knowledge retrieval.
  • Build scalable backend services using Python/FastAPI with type-safe APIs, authentication, and robust error handling.
  • Develop responsive, performant frontend applications using React/Next.js with real-time streaming for LLM responses.
  • Design and optimize database schemas across PostgreSQL, MongoDB, and Redis to support high-throughput AI workloads.
  • Implement WebSocket servers and event-driven architectures for real-time user experiences.
  • Create comprehensive testing strategies covering unit, integration, and end-to-end tests.
  • Deploy and manage ML/AI services using Docker containers and Kubernetes orchestration.
  • Build and maintain CI/CD pipelines for rapid, safe deployment of AI features.
  • Implement infrastructure as code using Terraform to manage cloud resources (AWS, Azure, or GCP).
  • Set up monitoring and observability using Datadog, Prometheus/Grafana, and LLM-specific tools (LangSmith, Weights & Biases).
  • Optimize costs through intelligent caching, batching strategies, and model selection algorithms.
  • Ensure enterprise-grade security through authentication, authorization, secrets management, and compliance measures.

Skills

Python
FastAPI
React
Kubernetes
Docker
LLM
RAG
Vector DB
Terraform
Cloud (AWS/Azure/GCP)

Tools

PostgreSQL
MySQL
MongoDB
Redis
Pinecone
Weaviate
Chroma
FAISS
Qdrant
Docker
Kubernetes
Terraform
AWS
Azure
GCP

Job description

Oteemo, a leading technology consulting firm in Reston, VA, seeks an exceptional AI Software Engineer to build and scale enterprise AI applications end to end, from database to UI. You will work with cutting-edge LLM tech, RAG systems, and production ML infrastructure, blending full-stack development with hands-on AI/ML engineering.

You will design end-to-end RAG pipelines, integrate with GPT-4/Claude/Gemini, and develop prompt strategies, agent systems, and scalable back-end/frontend services.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Systems Engineer — Production ML & RAG Expert
AI Systems Engineer — Production ML & RAG Expert

Oteemo, Inc • Reston (VA)

Hybrid
USD 140,000 - 200,000
AI Software Engineer - End-to-End Production ML & RAG
AI Software Engineer - End-to-End Production ML & RAG

Oteemo Inc. • Reston (VA)

On-site
USD 140,000 - 190,000
AI Systems Engineer — Production GenAI & RAG Pipelines
AI Systems Engineer — Production GenAI & RAG Pipelines

Oteemo, Inc • Virginia (IL)

On-site
USD 120,000 - 210,000
AI Software Engineer
AI Software Engineer

Oteemo, Inc • Reston (VA)

Hybrid
USD 140,000 - 200,000
AI Software Engineer
AI Software Engineer

Oteemo, Inc • Virginia (IL)

On-site
USD 120,000 - 210,000
AI Software Engineer
AI Software Engineer

Oteemo • Reston (VA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Production AI Engineer: LLM Systems & RAG Pipelines
Production AI Engineer: LLM Systems & RAG Pipelines

Elios AI • Denver (CO)

Hybrid
USD 140,000 - 190,000
Competitive compensation and benefits
AI Software Engineer with Security Clearance
AI Software Engineer with Security Clearance

Oteemo Inc. • Reston (VA)

On-site
USD 140,000 - 190,000
Senior AI/ML Full-Stack Engineer: RAG & LLM Orchestration
Senior AI/ML Full-Stack Engineer: RAG & LLM Orchestration

MAS Global Consulting • Plano (TX)

On-site
USD 140,000 - 190,000
Remote AI/ML Engineer — LLMs, RAG & Multi-Agent
Remote AI/ML Engineer — LLMs, RAG & Multi-Agent

YO AI Labs • Phoenix (AZ)

Remote
USD 120,000 - 180,000