AI/LLM Engineer — Healthcare Payment Integrity

Codoxo

Atlanta (GA)

Hybrid

USD 120,000 - 190,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Health insurance
Unlimited PTO
Annual home office stipend
401K match

Job summary

Codoxo is seeking an AI/LLM Engineer to design, ship, and operate LLM-powered features that accelerate claim audits and SIU investigations. You’ll handle retrieval, extraction, and summarization over sensitive documents, and build evaluation systems to prove effectiveness.

Our stack is Python and Django on AWS, using Bedrock for inference, OpenSearch for vector search, and Celery for processing. You will own features end to end and influence future multi-agent orchestration.

Qualifications

  • 3+ years of professional Python experience.
  • 2+ years building with LLMs in production.
  • Hands-on RAG experience: embeddings, vector or hybrid search.
  • Real practice around evaluation: know your eval set and metrics.
  • Working knowledge of AWS serverless: Lambda, S3, API Gateway; Bedrock experience is a plus.
  • Comfort with PHI or regulated data.

Responsibilities

  • Build and ship LLM features for investigator workflows (QA, extraction, summarization).
  • Own end-to-end retrieval pipeline: ingestion, extraction, chunking, filtering, hybrid search, citations.
  • Develop and maintain evaluation systems: golden sets, regression suites, rubric-based grading.
  • Treat prompts as engineering artifacts: versioned, code-reviewed, documented, with eval runs.
  • Design constrained schemas for structured output and handle refusals gracefully.
  • Work on serverless AWS stack (Bedrock, Lambda, S3, API Gateway, Step Functions).
  • Optimize cost and latency: token accounting, caching, routing tasks to model tiers.
  • Instrument features for observability: versions, token counts, retry rates, document IDs.
  • Collaborate with investigators, coders, and product to define eval criteria.

Skills

Python
LLMs in production
RAG and embeddings
Evaluation methodologies
AWS serverless
PHI handling

Tools

OpenSearch
Bedrock
AWS Lambda
Django

Job description

The United States spends roughly $4.9 trillion on healthcare each year, and an estimated quarter of that is lost to waste, fraud, abuse, and error. Codoxo is the premier provider of AI-driven solutions that help healthcare companies and government agencies proactively detect and reduce those losses and ensure payment integrity.

We are purpose-driven, with the goal of making healthcare more affordable and accessible to all. If you are passionate about applied AI and driven by positive impact in healthcare, Codoxo is where you belong. We are venture backed by some of the top investors in the country, with strong financials, and remain one of the fastest growing healthcare AI companies in the industry.

Position Summary

As an AI/LLM Engineer, you will design, ship, and operate LLM-powered features that accelerate claim audits and SIU investigations — retrieval and question answering over case documents, structured extraction from claims and medical records, and investigator-facing summarization. Just as importantly, you will build the evaluation systems that tell us whether any of it actually works.

Our production stack is Python and Django on AWS, with Amazon Bedrock for inference, OpenSearch for vector and hybrid search, and Celery for asynchronous document processing. You will work on a small team where you own features end to end and your work reaches fraud investigators at national health plans and government agencies.

As our LLM foundation matures, this role grows into multi-step, tool-using systems — MCP tools, Bedrock AgentCore, and multi-agent orchestration. You would help decide when that complexity is earned rather than inheriting the decision. We are looking for someone who reaches for the simplest thing that clears the bar and can explain why.

Key Responsibilities
  • Build and ship LLM features against real investigator workflows: document question answering, structured extraction, and summarization over claims, medical records, notes, and correspondence.
  • Own our retrieval pipeline end to end — document ingestion and extraction quality, chunking, metadata filtering, hybrid lexical and semantic search, reranking, and citations traced back to source text.
  • Build and maintain the evaluation systems that gate our releases: curated golden sets built with subject-matter experts, regression suites that run on every prompt and model change, and rubric-based grading validated against human reviewers.
  • Treat prompts as engineering artifacts — versioned, code-reviewed, documented with change notes, and never shipped without an evaluation run.
  • Design reliable structured output: schema-constrained generation, validation with bounded retries, and graceful handling of refusals, truncation, and partial results.
  • Build on serverless AWS — Bedrock, Lambda, S3, API Gateway, Step Functions — keeping PHI inside our trust boundary and out of logs and traces.
  • Keep cost and latency predictable as usage grows: token accounting, prompt caching, routing tasks to the right model tier, and asynchronous processing for non-interactive work.
  • Instrument LLM features for observability and audit — prompt and model versions, token counts, validation and retry rates, retrieved document IDs.
  • Partner with investigators, clinical coders, and product to turn expert judgment into evaluation criteria, not just requirements.
Qualifications
  • 3+ years of professional Python experience.
  • 2+ years building with LLMs in production — not prototypes or demos. You have shipped something real, watched it fail in ways you did not expect, and fixed it.
  • Hands-on RAG experience: embeddings, vector or hybrid search (OpenSearch, pgvector, FAISS, or similar), chunking strategy, and the judgment to diagnose whether a bad answer came from retrieval or generation.
  • A real practice around evaluation. You can describe how you knew a feature was working, what your eval set looked like, and a time a metric misled you.
  • Working knowledge of AWS serverless — Lambda, S3, API Gateway — and at least one managed LLM service. Bedrock experience is a plus, but strong OpenAI, Azure OpenAI, or Vertex AI experience transfers fine.
  • Comfort working with PHI or other regulated data, or genuine interest in learning to do it properly.

If you meet most of this list, we would rather see your application than not. We are more interested in how you reason about LLM systems than in whether you have used our exact stack.

Bonus Points
  • Healthcare domain data: FHIR, ICD-10, CPT, HCPCS, and payer claims experience. Valuable, and something we can teach the right engineer.
  • Document extraction and OCR pipelines — the unglamorous work that determines whether everything downstream succeeds.
  • Amazon Bedrock AgentCore (Memory, Gateway, Runtime), Strands Agents, MCP servers, or A2A.
  • Observability practice: CloudWatch, OpenTelemetry, structured logging, incident response.
  • Parameter-Efficient Fine-Tuning (LoRA/QLoRA) and knowing when it is the wrong tool.
  • Agent frameworks: LangGraph, LlamaIndex, CrewAI, or equivalent.
  • Kubernetes/EKS, GPU inference, and caching strategy.
  • Explainability, bias monitoring, and responsible-AI controls.
Physical Requirements

Work is performed in an office environment (either in our office or work-from-home) and requires the ability to work on a computer, operate standard office equipment, and work at a desk.

Benefits for You
  • Health, dental, and vision insurance with 100% employee premium coverage (starts day 1).
  • Unlimited PTO.
  • Annual home office stipend.
  • 401K match (after 90 days).
Accessibility Notice

If you need reasonable accommodation for any part of the employment process due to a physical or mental disability, please send an email to careers@codoxo.com with the subject “Accommodation”. Reasonable accommodation requests will be considered on a case-by-case basis.

We Are an Equal Opportunity Employer

Codoxo prohibits discrimination of any type and affords equal employment opportunities to employees and applicants without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by law. This policy applies to all terms and conditions of employment.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI/LLM Engineer — Healthcare Payment Integrity
AI/LLM Engineer — Healthcare Payment Integrity

Codoxo • Duluth (GA)

On-site
USD 130,000 - 190,000
Health, dental, and vision insurance
Unlimited PTO
Annual home office stipend
+1
Data Mining Auditor
Data Mining Auditor

Codoxo • United States

Hybrid
USD 85,000 - 125,000
Health insurance
Unlimited PTO
Professional development stipend
+2
Software Support Analyst
Software Support Analyst

Codoxo • United States

On-site
USD 52,000 - 65,000
Health, Dental, Vision insurance
Unlimited PTO
Professional development stipend
+2
AI Engineer
AI Engineer

Volo Health, LLC • Town of Florida (NY)

On-site
USD 120,000 - 190,000
Health insurance
Dental insurance
Vision insurance
+4
Senior Applied AI Engineer
Senior Applied AI Engineer

AgentGraph • Northern (KY)

On-site
USD 150,000 - 190,000
401(k) match
Company HSA contributions
Wellness reimbursement
+6
AI Engineer
AI Engineer

Latitude • Town of Florida (NY)

On-site
USD 120,000 - 180,000
Relocation assistance
Competitive salary
Health insurance
+1
Senior Software Engineer
Senior Software Engineer

UnitedHealth Group • Eden Prairie (MN)

On-site
Confidential
Comprehensive benefits package
Incentive programs
401k contribution
LLM Inference Engineer (Mid, Sr, Staff)
LLM Inference Engineer (Mid, Sr, Staff)

Hippocratic AI Inc. • Menlo Park (CA)

On-site
USD 180,000 - 280,000
Equity
Health insurance
Competitive compensation
Senior AI Engineer
Senior AI Engineer

CCC Intelligent Solutions Inc. • Chicago (IL)

On-site
USD 120,000 - 160,000
401K Match
Paid time off
Performance Bonus
+2
AI Enablement Engineer
AI Enablement Engineer

MediQuant • United States

Remote
USD 140,000 - 180,000