AI Application Engineer (Part-Time)

Workana LLC

United States

Remote

USD 70,000 - 120,000

Part time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Fully remote role
Real AI product work in legal/health
High ownership and autonomy
Performance-based incentives
Opportunity to transition to full-time

Job summary

Medxprts.ai is hiring an AI Engineer - LLM Fine-Tuning to hands-on train and fine-tune open-weight models, build training infra, and create robust evaluation and feedback pipelines. You will work with PyTorch, Hugging Face tooling and cloud providers to ship models into production.

Role requires collaboration with engineering teams, experience with LoRA/QLoRA and SFT, and overlap with U.S. time zones. Remote work with potential full-time transition is offered.

Qualifications

  • Hands-on experience fine-tuning LLMs, open-weight models and shipped example.
  • Experience building training infra with PyTorch & Hugging Face toolings.
  • Dataset engineering for messy unstructured docs; deduplication and train/eval splits.
  • Strong evaluation discipline with human-calibrated or regression testing.
  • Experience with preference/feedback learning methods and model updates.

Responsibilities

  • Design, run LLM fine-tuning experiments (LoRA/QLoRA and SFT) on open-weight models.
  • Build and maintain training infra across multi-GPU setups (DeepSpeed/FSDP).
  • Engineer domain datasets from long PDFs and medical/legal records.
  • Create evaluation harnesses and automated drift monitoring.
  • Integrate feedback loops (DPO/RLHF-style) into retraining pipelines.
  • Collaborate with backend/frontend teams to deploy models in production.

Skills

LLM fine-tuning
Dataset engineering
English communication

Tools

PyTorch
Hugging Face (Transformers)
PEFT
TRL/Axolotl
DeepSpeed/FSDP
AWS/GCP
Docker
Kubernetes

Job description

Client: medxprts.ai
Location: Remote
Type: Part-time, with potential to transition to Full-time
Schedule: U.S. timezone overlap required

Description

Medxprts.ai is building an AI-powered platform for the legal and healthcare space, using LLMs, agentic workflows, and automation to create production-grade applications.

We are seeking an AI Engineer - LLM Fine-Tuning: a hands-on engineer who has personally trained and fine-tuned open-weight models, built training infrastructure, engineered datasets from messy documents, and established rigorous evaluation and preference/feedback training pipelines. This role works closely with engineering teams to ship models and integrated features into production.

Responsibilities

Design, implement, and run LLM fine-tuning experiments (LoRA/QLoRA and full SFT) on open-weight models (e.g., Llama, Mistral, Qwen) and ship trained models into product workflows. Build and maintain training infrastructure using PyTorch and Hugging Face tooling (Transformers, PEFT, TRL/Axolotl), including multi-GPU training orchestration (DeepSpeed/FSDP) on AWS or GCP. Engineer datasets from real-world unstructured sources (long PDFs, medical/legal records), performing deduplication, filtering, contamination checks, and train/eval splits. Create evaluation harnesses tailored to domain needs: held-out test sets, LLM-as-judge with human calibration, regression tests across model versions, and automated monitoring for model drift. Implement preference/feedback training workflows (DPO/RLHF-style or similar) to learn from expert corrections (doctor-in-the-loop), and integrate feedback loops into model retraining pipelines. Collaborate with backend/frontend engineers to integrate fine-tuned models into services, optimize inference latency/cost, and support production deployments. Participate in PR reviews, release processes, incident debugging, and continuous improvement of training and deployment tooling. Ensure secure, compliant handling of sensitive data (HIPAA-awareness is highly preferred) during dataset preparation and model training.

Requirements

Hands-on experience fine-tuning LLMs: personally trained or fine-tuned open-weight models using LoRA/QLoRA and full SFT; able to explain trade-offs and provide at least one shipped example. Training infrastructure experience: PyTorch + Hugging Face ecosystem (Transformers, PEFT, TRL/Axolotl), multi-GPU training knowledge (DeepSpeed or FSDP), and running training workloads on AWS or GCP. Dataset engineering expertise: built instruction/preference datasets from messy, unstructured documents; practical knowledge of deduplication, filtering, train/eval splits, and contamination prevention. Strong evaluation discipline: designed domain-specific evaluation harnesses beyond standard benchmarks, including human-calibrated judge setups and regression testing. Practical experience with preference/feedback learning methods (DPO, RLHF-style workflows, or equivalent) and integrating expert feedback into model updates. Solid software engineering fundamentals: production workflows (Git, PRs), debugging, testing, deployment experience, and maintainable code. Experience with APIs, databases, and service integration for model inference. Ability to work independently, learn quickly, and follow technical direction. Good English communication skills and availability to overlap with U.S. working hours.
Nice to Have Experience with long-context handling strategies for very large documents (retrieval-aware training, context extension, RAG for multi-thousand-page sources). Model deployment and inference optimization: quantization (GPTQ/AWQ), vLLM/TGI serving, batching/throughput tuning, latency and cost optimization. Familiarity with vector databases, advanced RAG pipelines, MCPs, or n8n-style workflow automation tools. Knowledge of Docker, CI/CD, Kubernetes/EKS, or serverless infrastructure. Prior experience in healthcare, legaltech, HIPAA-aware processes, or other regulated/data-sensitive environments. Public portfolio, GitHub, or examples of shipped LLM/agentic applications and fine-tuning projects.

Benefits
  • Fully remote role.
  • Opportunity to work on real AI products in the legal and healthcare domain.
  • High ownership and autonomy.
  • Performance-based incentives and outcome-driven bonuses.
  • Potential to grow into a long-term, full-time role.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Engineer
AI Engineer

AllyNd Partners • Chicago (IL)

On-site
USD 120,000 - 160,000
Annual learning stipend
Flexible hours
Comprehensive health benefits
+1
Remote AI Engineer: LLM Fine-Tuning for Production
Remote AI Engineer: LLM Fine-Tuning for Production

Workana LLC • United States

Remote
USD 70,000 - 120,000
Fully remote role
Real AI product work in legal/health
High ownership and autonomy
+2
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Syndesus, Inc. • Austin (TX)

On-site
USD 100,000 - 140,000
100% employer-paid health, vision, and dental insurance
Retirement plans (401(k))
Disability insurance
+1
AI Engineer - TX
AI Engineer - TX

LawPro.ai • Town of Texas (WI)

On-site
USD 140,000 - 210,000
AI Engineer
AI Engineer

LawPro.ai, Inc. • Northern (KY)

Hybrid
USD 150,000 - 170,000
Equity stake
Unlimited PTO
Health, dental, vision benefits
+1
AI Engineer - VA
AI Engineer - VA

LawPro.ai • Virginia (MN)

On-site
USD 140,000 - 200,000
AI Engineer - NC
AI Engineer - NC

LawPro.ai • North Carolina

On-site
USD 140,000 - 190,000
AI Engineer - FL
AI Engineer - FL

LawPro.ai • Town of Florida (NY)

On-site
USD 140,000 - 210,000
LLM Training & Model Development Engineer
LLM Training & Model Development Engineer

InOpTra Digital • United States

On-site
USD 90,000 - 120,000
Competitive salary
Opportunity for remote work
Health benefits
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Syndesus, Inc. • Austin (TX)

On-site
USD 140,000 - 160,000
Health, vision, dental insurance (100%
401k retirement plans
Disability insurance
+1