Senior Machine Learning Engineer Chicago, IL

Attain

Chicago, Northern (IL, KY)

Hybrid

USD 170,000 - 240,000

Full time

1 hour ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Attain is seeking a Senior Machine Learning Engineer to own our production ML systems and build the MLOps platform infrastructure powering our B2C financial services. This role is hands-on and infrastructure-first, focusing on pipelines, feature infrastructure, model-serving, CI/CD, and observability in production.

You will design, build, and operate the end-to-end ML pipelines and tooling, enabling data scientists to deploy and retrain models quickly and safely while keeping systems healthy and

Qualifications

  • 5+ years of direct experience as a Machine Learning Engineer, ML Platform Engineer, MLOps Engineer, Applied Scientist or similar role building and operating production ML systems
  • Degree in STEM field such as Computer Science, Statistics, Economics, Mathematics, Engineering, Physics, Operations Research, or related quantitative field
  • Strong expertise deploying, serving, monitoring, and operating ML models in production—including feature engineering systems, training/serving parity, retraining, and model performance diagnostics
  • Experience building low-latency online model serving for real-time decisioning
  • Hands-on MLOps experience: pipelines, CI/CD for ML, containerization (Docker), orchestration (Kubernetes), infrastructure-as-code (Terraform), and workflow schedulers (Airflow)
  • Experience with model versioning, reproducibility, and safe progressive rollout of models in production
  • Fluently directing AI coding agents to build, operate, and debug production ML systems
  • Track record of replacing manual ML workflows with durable automation
  • Experience building infrastructure behind high-impact ML use cases such as credit decisioning, risk modeling, fraud, churn
  • Familiarity with model explainability and auditing for regulated decisioning
  • Strong software and platform engineering fundamentals
  • Strong Python coding skills; Go or Rust is a plus
  • Experience with distributed computing and GPU-accelerated workloads (Spark, Ray, Dask)
  • Strong SQL skills and cloud data warehouses (BigQuery, Spanner)
  • Experience with observability tools (Prometheus, Grafana, Datadog)
  • Experience with cloud platforms; GCP preferred
  • Willingness to wear multiple hats across engineering, infra, and ML execution
  • Strong written and verbal communication skills

Responsibilities

  • Build, deploy, and operate production ML systems focusing on reliability, performance, and fast execution
  • Improve pipelines and serving infrastructure behind predictive models across decisioning, fraud, churn, and transaction intelligence
  • Own production model lifecycle: features, deployment, CI/CD, monitoring, retraining
  • Develop reusable modeling pipelines, feature engineering systems, model-serving infra, and production-quality code deployed via Terraform and CI/CD in GCP + Kubernetes
  • Instrument models with monitoring, alerting, and retraining; define metrics/dashboards to surface drift
  • Direct AI coding agents to write, test, and ship infra and pipeline code with prudent verification
  • Automate repetitive ML lifecycle steps to speed up development
  • Collaborate with data scientists to deploy, iterate, and retrain models in production
  • Partner with analysts, platform engineers, product managers, and stakeholders to deliver ML systems with quality and efficiency
  • Identify areas for platform improvements, automation, and MLOps tooling to boost velocity

Skills

Python
MLOps
CI/CD for ML
SQL
Distributed computing
Cloud data warehouses
Prometheus/Grafana
Go or Rust

Education

B.S. in STEM field

Tools

Docker
Kubernetes
Terraform
Airflow
GCP
BigQuery
Istio
Prometheus/Grafana

Job description

Built for consumers and companies, alike.

Klover's engineering team powers one of the fastest-growing fintech platforms in the U.S., supporting over one million active users each month. Our systems process and move more than $1.5 billion annually, enabling real-time access to financial tools, rewards, and services that help people improve their day-to-day lives.

As part of this team, you'll help design, build, and scale the systems that underpin Klover's core products and platform. You'll work on high-impact, production-grade systems that prioritize reliability, security, and performance, and that integrate with a broad ecosystem of internal and external services. The work you do will directly shape how users interact with Klover's products, access their money, and experience transparent, low-fee financial services.

Klover engineers collaborate closely with colleagues across backend, frontend, data science, and product teams to deliver scalable, high-quality solutions for a rapidly growing user base. You'll have the opportunity to work with modern technologies and architectures while helping define and evolve the next generation of inclusive, data-powered financial products—building systems and interfaces that emphasize reliability, privacy, and performance at scale.

About the role

Attain is seeking a Senior Machine Learning Engineer to own our production ML systems and build out the MLOps platform infrastructure that powers our suite of B2C financial services. This role will be highly hands‑on and infrastructure‑first, focused on designing, building, and operating the pipelines, platforms, and tooling that take models from experiment to reliable production service across our app portfolio—and on keeping those systems healthy, performant, and cost‑effective once they're live.

You will work on the systems and infrastructure behind our high-impact predictive models, including the pipelines, feature infrastructure, model‑serving, CI/CD, and observability that keep them reproducible, automated, monitored, and fast in production. Day to day, this means building the platform and automation that let us move fast without sacrificing performance—streamlining retraining and rollouts, tuning systems for speed and efficiency, and building the metrics and alerting that give us confidence to ship—while enabling data scientists to deploy and iterate on models quickly and safely. The ideal candidate combines strong software and platform engineering fundamentals with practical MLOps experience building and operating production ML systems from scratch, and treats modern AI tooling as a first‑class part of how the work gets done—directing coding agents to write, test, and ship infrastructure code, with the judgment to know when to verify their work.

  • Chicago, IL: 4 days in-office; 1 day remote
What a typical week might look like
  • Build, deploy, and operate the production ML systems at the core of our EWA product, with a focus on reliability, performance, and fast, high-quality execution
  • Build and improve the pipelines and serving infrastructure behind our predictive models across consumer decisioning, fraud, churn, transaction intelligence, and other business-critical use cases
  • Own the production side of the model lifecycle: feature pipelines, deployment, CI/CD, monitoring, and automated retraining
  • Build and maintain reusable modeling pipelines, feature engineering systems, model‑serving infrastructure, and production-quality code, deployed via Terraform and CI/CD into our GCP + Kubernetes environment
  • Instrument models and pipelines with monitoring, alerting, and automated retraining—defining the metrics and dashboards (e.g., Prometheus/Grafana) that surface drift and degradation and give us confidence to ship
  • Direct AI coding agents as a force multiplier to write, test, and ship infrastructure and pipeline code—and apply strong judgment about when to trust their output and when to verify it yourself
  • Automate manual, repetitive steps in the ML lifecycle so the team can move faster without sacrificing reliability
  • Partner with data scientists to give them fast, safe paths to deploy, iterate on, and retrain models in production
  • Collaborate with analysts, platform engineers, product managers, and business stakeholders to deliver ML systems with quality, efficiency, and precision
  • Identify new areas where platform improvements, automation, and MLOps tooling can improve product velocity and business outcomes
Preferred Qualifications
  • 5+ years of direct experience as a Machine Learning Engineer, ML Platform Engineer, MLOps Engineer, Applied Scientist or similar role building and operating production ML systems
  • Strongly preferred: degree in STEM field such as Computer Science, Statistics, Economics, Mathematics, Engineering, Physics, Operations Research, or a related quantitative field
  • Demonstrated ability to apply critical thinking, abstract reasoning, and sound engineering judgment to complex, ambiguous technical and business problems
  • Strong expertise deploying, serving, monitoring, and operating ML models in production—including feature engineering systems, training/serving parity, retraining, and model performance diagnostics
  • Experience building low‑latency online model serving (e.g., gRPC/microservices, ideally with a service mesh such as Istio) for real‑time decisioning
  • Hands‑on MLOps experience: pipelines, CI/CD for ML, containerization (Docker), orchestration (Kubernetes), infrastructure‑as‑code (e.g., Terraform), and workflow schedulers (e.g., Airflow)
  • Experience with model versioning, reproducibility, and safe progressive rollout (shadow, canary, champion‑challenger) of models in production
  • Demonstrated fluency directing AI coding agents (e.g., Claude Code, Cursor, or similar) to build, operate, and debug real ML systems—with experienced judgment on verifying their work
  • A track record of replacing manual, repetitive ML workflows with durable automation
  • Experience building the infrastructure behind high‑impact applied ML use cases such as credit decisioning, risk modeling, fraud, churn, or consumer behaviour modeling
  • Familiarity with model explainability, auditability, and the compliance considerations of regulated decisioning (a plus for credit/fintech contexts)
  • Strong software and platform engineering fundamentals
  • Strong Python coding skills, with the ability to build pipelines, services, and production-quality tooling from scratch; experience with a systems or backend language such as Go or Rust is a plus
  • Experience with distributed computing and GPU‑accelerated workloads (e.g., Spark, Ray, Dask, or distributed training/inference), including scaling data and model pipelines across clusters
  • Strong SQL skills and experience with cloud data warehouses and operational databases (e.g., BigQuery, Spanner), including working with large, messy, real‑world datasets
  • Experience with observability tools such as Prometheus, Grafana, or Datadog
  • Experience with cloud computing services or platforms; GCP preferred
  • Willingness to roll up your sleeves and wear multiple hats across engineering, infrastructure, and ML execution based on business needs
  • Strong written and verbal communication skills, including the ability to explain technical topics to both technical and non‑technical audiences
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Machine Learning Engineer
Senior Machine Learning Engineer

Socket.dev • Chicago (IL)

On-site
USD 140,000 - 190,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Attain • Chicago (IL)

On-site
USD 150,000 - 210,000
Senior Machine Learning Engineer
Senior Machine Learning Engineer

ExaCare AI • New York (NY)

On-site
USD 100,000 - 140,000
Flexible PTO
Medical, dental, and vision coverage
Company off-sites
Senior ML Engineer – MLOps & Mechanistic Interpretability
Senior ML Engineer – MLOps & Mechanistic Interpretability

Equifax • Alpharetta (GA)

Hybrid
USD 140,000 - 190,000
Machine Learning Engineer
Machine Learning Engineer

Darwill • Oak Brook (IL)

Hybrid
USD 120,000 - 160,000
Machine Learning Engineer
Machine Learning Engineer

Darwill • Northern (KY)

Hybrid
USD 120,000 - 160,000
Hybrid work model
Machine Learning Engineer
Machine Learning Engineer

Darwill, Inc. • Oakbrook Terrace (IL)

On-site
USD 120,000 - 180,000
Machine Learning Engineer
Machine Learning Engineer

Socket.dev • Shelton (CT)

On-site
USD 140,000 - 190,000
MLOps Engineer
MLOps Engineer

Compunnel, Inc. • San Antonio (TX)

On-site
USD 100,000 - 130,000
Senior ML Engineer – MLOps & Mechanistic Interpretability
Senior ML Engineer – MLOps & Mechanistic Interpretability

Equifax, Inc. • Alpharetta (GA)

On-site
USD 140,000 - 210,000
4 days in-office collaboration (Mon-TH
Friday Flexibility