Senior ML Engineer with Python (IR-536)

Intellectsoft

United States

Remote

USD 140,000 - 230,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Impactful projects
Udemy courses
Team events
Training workshops
Career path
Absence days
Flexible hours
Work from anywhere

Job summary

Intellectsoft is seeking a senior AI/ML backend engineer to design and build scalable ML platforms and production-grade APIs. You will integrate LLMs, implement multi-tenant architectures, and advance model serving with robust observability.

The role emphasizes Python, FastAPI, LangChain workflows, and end-to-end ML pipelines across NLP, CV, OCR, and deployment scenarios.

Qualifications

  • Bachelor's or Master’s degree in Computer Science or a related field.
  • Strong Python coding skills and ML/LLM experience.
  • Production ML/LLM systems exposure.

Responsibilities

  • Build scalable backend systems and APIs (FastAPI).
  • Design MLOps with KPI tracking, drift detection, and feedback loops.
  • Deploy and operationalize ML/Deep Learning models (LLMs).
  • Integrate Azure OpenAI and other LLMs with robust retry logic.

Skills

Python
ML/LLM systems
Async programming
FastAPI
SQLAlchemy
Dependency injection
Interfaces
Abstract base classes
Vector databases
Pinecone
Weaviate
Chroma
RAG architectures
LangChain
LangGraph
Pydantic
SOLID principles
MLflow
Model versioning
A/B testing
Langfuse
NLP
Computer vision
OCR
GPT-4 Vision
Feature pipelines
Real-time inference
Batch inference
Model serving
Hugging Face
LlamaIndex
SQL
Problem-solving
Team collaboration

Education

Bachelor’s or Master’s degree in Computer Science or a related field

Tools

FastAPI
SQLAlchemy
Pinecone
Weaviate
Chroma
LangChain
LangGraph
MLflow
Langfuse
Azure OpenAI
Hugging Face
LlamaIndex

Job description

Our customer’s product is an AI-powered platform that helps businesses make better decisions and work more efficiently. It uses advanced analytics and machine learning to analyze large amounts of data and provide useful insights and predictions. The platform is widely used in various industries, including healthcare, to optimize processes, improve customer experiences, and support innovation. It integrates easily with existing systems, making it easier for teams to make quick, data-driven decisions to deliver cutting-edge solutions.

  • Bachelor’s or Master’s degree in Computer Science or a related field.
  • Strong Python coding skills - 7+ years.
  • 2+ years of hands-on experience with machine learning and production LLM systems.
  • Experience building backend APIs with FastAPI, async patterns, rate limiting, and SQLAlchemy - 3+ years.
  • Experience designing maintainable and extensible systems using dependency injection, interfaces, and abstract base classes.
  • Experience with vector databases such as Pinecone, Weaviate, or Chroma, as well as hybrid search.
  • Strong understanding of RAG architectures, including retrieval, reranking, context assembly, and response generation.
  • Hands-on experience with LangChain and LangGraph for building and orchestrating LLM workflows.
  • Advanced Python skills, including async/await, type hints, Pydantic, and SOLID principles.
  • MLOps experience with MLflow, model versioning, and A/B testing; experience with Langfuse is a plus.
  • Experience in NLP and computer vision, including document understanding, OCR, and GPT-4 Vision.
  • Experience building feature pipelines, real-time and batch inference systems, and model serving.
  • Hands-on experience with Hugging Face is required; experience with LlamaIndex is a plus.
  • Familiarity with database technologies such as SQL.
  • Good problem-solving skills and the ability to work in a fast-paced, team-oriented environment.
Nice to have skills:
  • Understanding of DevOps, CI / CD including: Docker containerization, Azure DevOps pipelines or GitHub Actions, Kubernetes (nice to have);
  • Data security including: Multi-tenant data isolation, Secure key management (Azure Key Vault), Audit trail implementation;
  • Experience in designing on cloud platform including: Azure (strongly preferred): Azure OpenAI, Blob Storage, Key Vault, Container Registry, AWS or GCP;
  • Experience in data engineering in Big Data systems including: Large-scale data processing, ETL/ELT pipelines.
  • Rate limiting and quota management for high-throughput API usage.
  • Cost management and optimization for LLM usage at scale.
  • Document processing expertise (PDF extraction, OCR tooling).
  • Production incident management and on-call experience.
  • Testing strategies for non-deterministic LLM outputs (e.g., golden datasets, fuzzy matching).
  • Domain knowledge in regulated industries (e.g., healthcare/pharma workflows, regulatory compliance) is a plus.
Responsibilities:
  • Build, refine, and use ML Engineering platforms and components; develop and implement scalable backend systems, APIs, and microservices using FastAPI.
  • Implement MLOps including model KPI measurement, tracking, model drift detection, and model feedback loops.
  • Deploy and operationalize ML and Deep Learning models, with a strong focus on LLMs and Generative AI.
  • Integrate Azure OpenAI (GPT-4, GPT-4 Vision) and other LLM providers with proper retry logic and error handling.
  • Maintain up-to-date knowledge of state-of-the-art technologies such as LLMs, GenAI, and transformer architectures.
  • Scale machine learning algorithms to work on massive data sets under strict SLAs.
  • Build and orchestrate model pipelines including feature engineering, inferencing, and continuous model training.
  • Write backend application code in Python and SQL using strong object-oriented principles and asynchronous programming (asyncio, async/await).
  • Implement dependency injection patterns and layered architecture (Service, Foundation, Orchestration, DAL).
  • Build LLM observability (e.g., Langfuse) to track prompts, tokens, costs, and latency.
  • Develop prompt management systems with versioning and fallback mechanisms.
  • Implement Celery (or similar) workflows for asynchronous task processing and complex pipelines.
  • Build multi-tenant architectures with client data isolation.
  • Implement cost optimization strategies for LLM usage (prompt caching, batch processing, token optimization).
  • Integrate third-party APIs and services (e.g., document/OCR services, cloud storage, enterprise systems).
  • Collaborate with client-facing teams to understand business context and contribute to technical requirement gathering.
  • Write production-ready code that is testable, maintainable, and accounts for edge cases and errors.
  • Ensure high quality of deliverables by following architecture/design guidelines, coding best practices, and periodic design/code reviews.
  • Write unit tests and higher-level tests to handle expected edge cases and errors gracefully.
  • Troubleshoot backend application code using structured logging and distributed tracing.
  • Use bug tracking, code review, version control, and other tools to organize and deliver work.
  • Participate in scrum calls and agile ceremonies, communicating progress, issues, and dependencies.
  • Document application changes and updates, including API documentation via OpenAPI/Swagger.
  • Research and evaluate emerging architecture patterns and technologies through rapid learning, proofs-of-concept, and prototypes.
  • Awesome projects with an impact
  • Udemy courses of your choice
  • Team-buildings, events, marathons & charity activities to connect and recharge
  • Workshops, trainings, expert knowledge-sharing that keep you growing
  • Clear career path
  • Absence days for work-life balance
  • Flexible hours & work setup - work from anywhere and organize your day your way
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Python Engineer with AI Exposure
Senior Python Engineer with AI Exposure

Aether Biomedical • United States

On-site
USD 140,000 - 190,000
Vacation days 26
Health insurance
Mental health support
+3
[Job-31573] AI Engineer Master
[Job-31573] AI Engineer Master

ciandt • United States

Hybrid
USD 180,000 - 270,000
Health insurance
Dental insurance
Life insurance
+8
Applied AI Engineer
Applied AI Engineer

SherlockTalent • Miami (FL)

On-site
USD 120,000 - 140,000
Solid Benefits
Referral bonus of $2,500
Senior ML Engineer (Applied AI)
Senior ML Engineer (Applied AI)

Internetwork Expert • Massachusetts

On-site
USD 140,000 - 210,000
Annual paid vacation
Health Insurance
Remote-first culture
+1
Principal Engineer
Principal Engineer

your Jared • Northern (KY), San Diego (CA)

Hybrid
USD 180,000 - 240,000
Fully remote, work from home
Employee Share Option Plan
Flexible working hours
+4
AI/ML Engineer
AI/ML Engineer

Winaxis LLC • Dallas (TX)

On-site
USD 120,000 - 160,000
Senior AI Engineer II
Senior AI Engineer II

DataJobs • Atlanta (GA)

On-site
USD 140,000 - 210,000
Bonus opportunities
401k with company match
Employee Stock Purchase Plan
+2
AIML Engineer
AIML Engineer

Qubeaxis • San Francisco (CA)

On-site
USD 180,000 - 260,000
Performance bonus (up to 20% of base)
Equity participation
Health, dental, and vision insurance
+3
Senior AI ML Engineer with OpenAI API, NLP Enrichment Platform
Senior AI ML Engineer with OpenAI API, NLP Enrichment Platform

Aether Biomedical • United States

On-site
USD 140,000 - 190,000
Vacation days up to 26 business days/​
Health and life insurance
Flexible workplace: office/home/hybrid
+2
AI Engineer
AI Engineer

Valsoft Corporation • Northern (KY)

On-site
USD 120,000 - 180,000