AI Engineer | Agentic Systems Machinify · Remote · US · AI Engineering $130,000–$200,000 7mo ago

Aimlroles

Northern (KY)

Hybrid

USD 130,000 - 200,000

Full time

14 days+
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Work from anywhere in the US
Top Medical/Dental/Vision offerings
FSA/HSA
Tuition reimbursement
Competitive salary with 401(k) match
Unlimited PTO
Flexible work environment

Job summary

Machinify is hiring an L4 AI Engineer to design agent systems from scratch, auditing medical claims end-to-end, and delivering defensible findings for regulatory review. You will decide on architecture, tool usage, and context strategy, shaping systems that operate with high accuracy and reliability in a complex domain.

Join a production-grade environment that emphasizes measurable results, deep domain knowledge in healthcare, and collaboration with a team that values engineering rigor and tool

Qualifications

  • 2-4 years of applied ML / AI engineering experience with a relevant degree or a Master's with no prior industry experience.
  • Strong Python engineering with clean abstractions and tested code.
  • Hands-on understanding of agent loops and tool integration.
  • Experience with major agent SDKs and pragmatic tradeoffs.
  • Ability to design context loading, prompts, and long-running tasks with grounding.

Responsibilities

  • Design agent systems from first principles and defend architecture choices.
  • Engineer the context, prompts, and tool surfaces for reliable outputs.
  • Drive evaluation rigor with pre-built evals and metric tracking.
  • Utilize Claude Code, Codex, and related tools to plan, scaffold, and debug work.
  • Become a domain expert in healthcare claims, coding guidelines, and clinical records.

Skills

Agent loops understanding
Python engineering
OpenAI / Claude Codex
VS Code & Git
Evaluation metrics

Education

Bachelor's degree in CS/Math/Engineering or equivalent

Tools

OpenAI Agents SDK
Claude Code / Codex
LangGraph

Job description

Machinify is a leading healthcare intelligence company with expertise across the payment continuum, delivering unmatched value, transparency, and efficiency to health plan clients across the country. Deployed by over 85 health plans, including many of the top 20, and representing more than 270 million lives, Machinify brings together a fully configurable and content-rich, AI-powered platform along with best-in-class expertise. We’re constantly reimagining what’s possible in our industry, creating disruptively simple, powerfully clear ways to maximize financial outcomes and drive down healthcare costs.

Machinify is a leading healthcare intelligence company with expertise across the payment continuum, delivering unmatched value, transparency, and efficiency to health plan clients across the country. Deployed by over 85 health plans - including many of the top 20 and representing more than 270 million lives - Machinify brings together a fully configurable, content-rich, AI-powered platform along with best-in-class expertise. We're constantly reimagining what's possible in our industry, creating disruptively simple, powerfully clear ways to maximize financial outcomes and drive down healthcare costs.

The Role

We're building production-grade agentic systems that audit medical claims end-to-end - reading raw medical records, reasoning over coding and clinical guidelines, and producing defensible findings that hold up to clinical and regulatory review. Reaching human-expert accuracy on noisy, long-context documents is one of the hardest unsolved problems in applied AI, and the field is moving weekly.

We're hiring an L4 AI Engineer who can step into an ambiguous problem, design an agent system from scratch, and ship it. You won't be plugging into someone else's architecture - you'll be deciding what the architecture should be.

What You'll Do
  • Design agent systems from first principles. Decide the loop, the tools, the context strategy, the evaluation harness. Choose between single-agent and multi-agent topologies, between LLM reasoning and deterministic post-passes, between retrieval and direct context loading - and defend the choice with data.
  • Engineer the context. The hardest part of building a good agent is what goes into the prompt and what comes out. You'll obsess over context windows, tool surfaces, structured outputs, citation grounding, and the prompt itself.
  • Drive evaluation rigor. Build evals before you build the agent. Diagnose where it fails, fix the root cause, and prove the fix moved the metric.
  • Use AI tooling like a power user. A meaningful fraction of your day will be spent driving Claude Code, Codex, and similar tools to plan, scaffold, refactor, and debug your own work. We expect you to be faster with these tools than most engineers are without them.
  • Become a domain expert. Healthcare claims, coding guidelines, and the medical record itself are unavoidable parts of the job. Strong engineers who lean into the domain become outsized contributors here.
What We're Looking For
Required
  • 2-4 years of applied ML / AI engineering experience with a Bachelor's in CS, Math, Engineering or equivalent - or a Master's in a similar program with no prior industry experience required. Either way, at least one production-quality system (industry, research, or substantial open-source) you owned end-to-end.
  • Strong Python engineering. Clean abstractions, type discipline, async, tested code.
  • Deep, hands-on understanding of agent loops - how a model decides to call a tool, how a tool result re-enters context, how loops terminate, where they fail.
  • Hands-on experience with at least one major agent SDK - OpenAI Agents SDK, Anthropic SDK / claude-agent-sdk, LangGraph, or equivalent - and an opinion on the tradeoffs.
  • Working knowledge of how modern coding agents are built and how they engineer context - what goes in the system prompt, how files are read and edited, how long-running tasks are planned and tracked, where they break.
  • Fluency with Claude Code / Codex as a power user. You should be able to brainstorm, plan, and execute non-trivial engineering tasks with these tools - including reading their source when needed to understand or extend behaviour.
  • Solid command of VS Code and git - branches, rebases, worktrees, conflict resolution, PR workflows. Not optional.
  • A bias toward measurement: you don't ship without an eval, and you don't believe a number you can't reproduce.
Strongly preferred
  • Experience designing structured outputs (Pydantic / JSON Schema) and tool interfaces that LLMs reliably call correctly.
  • Familiarity with reasoning models (o-series, Claude extended thinking, Gemini thinking) and a sense of when they earn their cost.
  • Prior work on long-context, citation-grounded systems where the model must point to evidence, not just answer.
  • Healthcare, legal, finance, or any other domain where "mostly right" is unacceptable.
Nice to have
  • Document understanding (OCR, layout-aware models, table extraction).
  • Vision-language models, multimodal retrieval.
  • Production experience with caching, observability, and cost control on LLM workloads.
What We Offer
  • Workfrom anywhere in the US!Machinifyis digital-first.
  • Top Medical/Dental/Vision offerings
  • FSA/HSA
  • Tuition reimbursement
  • Competitive salary, 401(k) with company match
  • Unlimited PTO
  • Additional health and wellness benefits and perks
  • Flexible and trusting environment where you’ll feel empowered to do your best work

The salary for this position is based on an array of factors unique to each candidate: Such as years and depth of experience, set skills, certifications, etc. We are hiring for different levels and the base salary can range from $130k-$200k+ based on your assessed level. Compensation also includes meaningful equity, healthcare, unlimited PTO, and more.

Equal Employment Opportunity at Machinify

We are committed to equal employment opportunity regardless of race, color, ancestry, religion, sex, national origin, sexual orientation, age, citizenship, marital status, disability, gender, gender identity or expression, or veteran status. We are proud to be an equal opportunity workplace. Machinify is an employment at will employer. We participate in E-Verify as required by applicable law. In accordance with applicable state laws, we do not inquire about salary history during the recruitment process. If you require a reasonable accommodation to complete any part of the application or recruitment process, please let our recruiters know. See our Candidate Privacy Notice at: https://www.machinify.com/candidate-privacy-notice/

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Software Engineer, Backend | Audit Product Team
Senior Software Engineer, Backend | Audit Product Team

The Rawlings Group • United States

On-site
USD 200,000 - 230,000
Work from anywhere in the US
Top Medical/Dental/Vision
FSA/HSA
+5
Sr/Staff Software Engineer, Backend | Web Products
Sr/Staff Software Engineer, Backend | Web Products

The Rawlings Group • United States

On-site
USD 200,000 - 250,000
Remote work in US
Competitive salary package
401(k) with company match
+3
Senior Data Engineer - Analytics New Remote - US
Senior Data Engineer - Analytics New Remote - US

Machinify • Northern (KY)

Hybrid
USD 170,000 - 200,000
Full Medical/Dental/Vision for you and
Family coverage
Flexible and trusting environment
+1
Remote Senior Backend Engineer — AI‑Driven Health Platform
Remote Senior Backend Engineer — AI‑Driven Health Platform

The Rawlings Group • United States

On-site
USD 200,000 - 230,000
Work from anywhere in the US
Top Medical/Dental/Vision
FSA/HSA
+5
Staff Data Scientist | NLP Machinify · Remote · US · AI Research $180,000–$230,000 7mo ago
Staff Data Scientist | NLP Machinify · Remote · US · AI Research $180,000–$230,000 7mo ago

Aimlroles • Northern (KY)

Hybrid
USD 180,000 - 230,000
Work from anywhere in the US
Top medical/dental/vision benefits
401(k) with company match
+2
AI Engineer (Remote)
AI Engineer (Remote)

M3 USA • United States

On-site
USD 120,000 - 180,000
Health and Dental
Life, Accident and DisabilityInsurance
Prescription Plan
+4
Machine Learning Engineer
Machine Learning Engineer

HealthEdge Software, Inc. • Northern (KY)

On-site
USD 90,000 - 135,000
Lead Medical Director - Claim Editing
Lead Medical Director - Claim Editing

Machinifyinc • United States

On-site
USD 260,000 - 380,000
Machine Learning Engineer
Machine Learning Engineer

HealthEdge • United States

On-site
USD 130,000 - 165,000
Senior Applied AI Engineer
Senior Applied AI Engineer

AgentGraph • Northern (KY)

On-site
USD 150,000 - 190,000
401(k) match
Company HSA contributions
Wellness reimbursement
+6