AI Software Engineer

Oteemo

Reston, Northern (VA, KY)

Hybrid

USD 140,000 - 190,000

Full time

6 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Oteemo, a leading technology consulting firm in Reston, VA, seeks an exceptional AI Software Engineer to build and scale enterprise AI applications end to end, from database to UI. You will work with cutting-edge LLM tech, RAG systems, and production ML infrastructure, blending full-stack development with hands-on AI/ML engineering.

You will design end-to-end RAG pipelines, integrate with GPT-4/Claude/Gemini, and develop prompt strategies, agent systems, and scalable back-end/frontend services.

Qualifications

  • Expert-level proficiency in Python with modern frameworks (FastAPI, Flask).
  • Solid understanding of relational (PostgreSQL, MySQL) and NoSQL (MongoDB) databases.
  • Experience with authentication systems (OAuth2, JWT, SSO) and security best practices.
  • Proven track record of shipping high-quality, scalable software to production.
  • Hands-on experience building and deploying AI/ML applications in production environments.
  • Deep understanding of LLM integration, prompt engineering, and context management.
  • Proven expertise with RAG systems, including document processing, chunking, embedding, retrieval, and generation.
  • Experience working with vector databases (Pinecone, Weaviate, Chroma, FAISS, or Qdrant).
  • Strong grasp of semantic search, similarity algorithms, and hybrid search techniques.
  • Knowledge of evaluation frameworks for assessing AI system quality and performance.
  • Production experience with Docker containerization and Kubernetes orchestration.
  • Strong knowledge of at least one major cloud platform (AWS, Azure, or GCP) and its AI services.
  • Experience building CI/CD pipelines for ML/AI application.
  • Proficiency with infrastructure as code tools (Terraform, CloudFormation, Pulumi).
  • Understanding of monitoring, logging, and alerting best practices; cost optimization experience for cloud and AI workloads.
  • Strong computer science fundamentals and algorithmic thinking.
  • Proficiency with Git workflows, code review practices, and collaborative development.
  • Excellent debugging and problem-solving skills.
  • Clear technical communication and documentation abilities.

Responsibilities

  • Design and implement end-to-end RAG pipelines for intelligent document search and question-answering across enterprise knowledge bases.
  • Build production-ready integrations with leading LLMs (GPT-4, Claude, Gemini) for accurate, contextual responses to user queries.
  • Develop prompt engineering strategies and evaluation frameworks to ensure consistent, high-quality AI outputs.
  • Create agent systems with tool integration capabilities that can autonomously complete complex tasks.
  • Implement vector search solutions using Pinecone, Weaviate, or similar technologies for semantic similarity and knowledge retrieval.
  • Build scalable backend services using Python/FastAPI with type-safe APIs, authentication, and robust error handling.
  • Develop responsive, performant frontend applications using React/Next.js with real-time streaming for LLM responses.
  • Design and optimize database schemas across PostgreSQL, MongoDB, and Redis to support high-throughput AI workloads.
  • Implement WebSocket servers and event-driven architectures for real-time user experiences.
  • Create comprehensive testing strategies covering unit, integration, and end-to-end tests.
  • Deploy and manage ML/AI services using Docker containers and Kubernetes orchestration.
  • Build and maintain CI/CD pipelines for rapid, safe deployment of AI features.
  • Implement infrastructure as code using Terraform to manage cloud resources (AWS, Azure, or GCP).
  • Set up monitoring and observability using Datadog, Prometheus/Grafana, and LLM-specific tools (LangSmith, Weights & Biases).
  • Optimize costs through intelligent caching, batching strategies, and model selection algorithms.
  • Ensure enterprise-grade security through authentication, authorization, secrets management, and compliance measures.

Skills

Python
FastAPI
React
Kubernetes
Docker
LLM
RAG
Vector DB
Terraform
Cloud (AWS/Azure/GCP)

Tools

PostgreSQL
MySQL
MongoDB
Redis
Pinecone
Weaviate
Chroma
FAISS
Qdrant
Docker
Kubernetes
Terraform
AWS
Azure
GCP

Job description

Oteemo is an industry-leading technology consulting firm at the forefront of cloud native, enterprise DevSecOps, and AI-driven transformation. We build intelligent, automated, secure systems for organizations tackling their toughest technical challenges and we're pushing the boundaries of what AI, generative AI, and agentic systems can do in production, not just in theory. Join us and you'll work alongside recognized experts on cutting-edge projects that blend cloud native architecture, extreme automation, and AI/ML at the core. We foster a dynamic, inclusive, and collaborative culture built on continuous learning, where your ideas shape real outcomes for our clients. If you're passionate about building what's next in AI and cloud technology and want to do it with a team that sets the standard rather than follows it, Oteemo is where you belong.

Job Description

We're seeking an exceptional AI Software Engineer to build and scale enterprise AI applications end to end from database to UI. In this role, you'll work with cutting-edge LLM technology, RAG systems, and production ML infrastructure, combining full-stack development expertise with hands‑on AI/ML engineering to ship intelligent systems that deliver real business value at scale. You'll be a key technical contributor, shipping production‑ready AI features that users love while ensuring reliability, performance, and cost‑effectiveness.

Key Responsibilities:

  • Design and implement end-to‑end RAG (Retrieval‑Augmented Generation) pipelines for intelligent document search and question‑answering across enterprise knowledge bases.
  • Build production‑ready integrations with leading LLMs (GPT‑4, Claude, Gemini) for accurate, contextual responses to user queries.
  • Develop prompt engineering strategies and evaluation frameworks to ensure consistent, high‑quality AI outputs.
  • Create agent systems with tool integration capabilities that can autonomously complete complex tasks.
  • Implement vector search solutions using Pinecone, Weaviate, or similar technologies for semantic similarity and knowledge retrieval.
  • Build scalable backend services using Python/FastAPI with type‑safe APIs, authentication, and robust error handling.
  • Develop responsive, performant frontend applications using React/Next.js with real‑time streaming for LLM responses.
  • Design and optimize database schemas across PostgreSQL, MongoDB, and Redis to support high‑throughput AI workloads.
  • Implement WebSocket servers and event‑driven architectures for real‑time user experiences.
  • Create comprehensive testing strategies covering unit, integration, and end‑to‑end tests.
  • Deploy and manage ML/AI services using Docker containers and Kubernetes orchestration.
  • Build and maintain CI/CD pipelines for rapid, safe deployment of AI features.
  • Implement infrastructure as code using Terraform to manage cloud resources (AWS, Azure, or GCP).
  • Set up monitoring and observability using Datadog, Prometheus/Grafana, and LLM‑specific tools (LangSmith, Weights & Biases).
  • Optimize costs through intelligent caching, batching strategies, and model selection algorithms.
  • Ensure enterprise‑grade security through authentication, authorization, secrets management, and compliance measures.
Qualifications
  • Expert‑level proficiency in Python with modern frameworks (FastAPI, Flask).
  • Solid understanding of relational (PostgreSQL, MySQL) and NoSQL (MongoDB) databases.
  • Experience with authentication systems (OAuth2, JWT, SSO) and security best practices.
  • Proven track record of shipping high‑quality, scalable software to production.
  • Hands‑on experience building and deploying AI/ML applications in production environments.
  • Deep understanding of LLM integration, prompt engineering, and context management.
  • Proven expertise with RAG systems, including document processing, chunking, embedding, retrieval, and generation.
  • Experience working with vector databases (Pinecone, Weaviate, Chroma, FAISS, or Qdrant).
  • Strong grasp of semantic search, similarity algorithms, and hybrid search techniques.
  • Knowledge of evaluation frameworks for assessing AI system quality and performance.
  • Production experience with Docker containerization and Kubernetes orchestration.
  • Strong knowledge of at least one major cloud platform (AWS, Azure, or GCP) and its AI services.
  • Experience building CI/CD pipelines for ML/AI application.
  • Proficiency with infrastructure as code tools (Terraform, CloudFormation, Pulumi).
  • Understanding of monitoring, logging, and alerting best practices; cost optimization experience for cloud and AI workloads.
  • Strong computer science fundamentals and algorithmic thinking.
  • Proficiency with Git workflows, code review practices, and collaborative development.
  • Excellent debugging and problem‑solving skills.
  • Clear technical communication and documentation abilities.
Additional Information

All your information will be kept confidential according to EEO guidelines.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Software Engineer
AI Software Engineer

Oteemo, Inc • Virginia (IL)

On-site
USD 120,000 - 210,000
AI Software Engineer
AI Software Engineer

Oteemo, Inc • Reston (VA)

Hybrid
USD 140,000 - 200,000
AI Software Engineer with Security Clearance
AI Software Engineer with Security Clearance

Oteemo Inc. • Reston (VA)

On-site
USD 140,000 - 190,000
Sr. Data Engineer
Sr. Data Engineer

Oteemo, Inc • Virginia (IL)

On-site
USD 140,000 - 200,000
Sr. Data Engineer with Security Clearance
Sr. Data Engineer with Security Clearance

Oteemo Inc. • Hampton (VA)

On-site
USD 120,000 - 180,000
Sr. Data Engineer
Sr. Data Engineer

Oteemo • United States

On-site
USD 120,000 - 160,000
AI Systems Engineer — Production ML & RAG Expert
AI Systems Engineer — Production ML & RAG Expert

Oteemo, Inc • Reston (VA)

Hybrid
USD 140,000 - 200,000
AI Software Engineer - End-to-End Production ML & RAG
AI Software Engineer - End-to-End Production ML & RAG

Oteemo Inc. • Reston (VA)

On-site
USD 140,000 - 190,000
Staff AI Software Engineer
Staff AI Software Engineer

Harnham • San Francisco (CA)

On-site
USD 150,000 - 200,000
AI Engineer
AI Engineer

LeoTechnologies • Boca Raton (FL), Northern (KY)

Hybrid
USD 140,000 - 190,000