MLOps Engineer

Evlo AI

Washington (District of Columbia)

On-site

USD 140,000 - 210,000

Full time

14 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Evlo AI is seeking an MLOps Engineer to own infrastructure and automation for productionizing ML and GenAI systems. You will build repeatable pipelines for training, evaluation, deployment, monitoring, and rollback across cloud platforms.

You will collaborate with ML engineers, data scientists, and platform teams to improve model release velocity while preserving security, reproducibility, and reliability. The role spans Kubernetes, CI/CD, registries, observability, and serving systems for

Qualifications

  • 3–8 years of experience in MLOps, platform engineering, DevOps, or software engineering
  • Strong Python and Linux skills
  • Hands-on experience with Docker, Kubernetes, infrastructure as code, and at least one major cloud platform
  • Practical knowledge of ML lifecycle tooling such as MLflow, Kubeflow, SageMaker, Vertex AI, Azure ML, Airflow, or comparable platforms
  • Experience designing CI/CD pipelines and observability using Prometheus, Grafana, OpenTelemetry, ELK, or cloud-native monitoring services
  • Equivalent professional experience considered when missing a degree

Responsibilities

  • Build and maintain ML pipelines for data validation, feature processing, model training, evaluation, registration, and promotion
  • Automate model and service deployment with Docker, Kubernetes, Helm, and Terraform across AWS, Azure, or GCP
  • Develop CI/CD workflows for Python services, model artifacts, infrastructure changes, and container images
  • Implement production monitoring for service health, latency, throughput, resource utilization, data drift, model performance, and quality regressions
  • Operate scalable model-serving infrastructure using platforms such as KServe, Seldon, NVIDIA Triton, Ray Serve, or managed cloud endpoints
  • Establish reproducibility and governance practices for datasets, features, model versions, experiments, secrets, and deployment approvals
  • Troubleshoot production incidents, improve system reliability, and document runbooks, architecture decisions, and operational standards

Skills

Python
Linux
Docker
Kubernetes
Cloud platforms
ML tooling
CI/CD
Observability

Education

Bachelor's degree in CS/Engineering/Math

Tools

Kubeflow
Airflow
MLflow
SageMaker
Vertex AI
Azure ML

Job description

About The Role

The MLOps Engineer owns the infrastructure and automation required to move machine learning and GenAI systems from experimentation into reliable production environments. The role builds repeatable pipelines for training, evaluation, deployment, monitoring, and rollback across cloud-based compute and data platforms.

You will work with ML engineers, data scientists, and platform teams to improve model release velocity without compromising security, reproducibility, or operational reliability. The work spans Kubernetes, CI/CD, model registries, observability, and the serving systems that support latency-sensitive AI products.

Key Responsibilities
  • Build and maintain ML pipelines for data validation, feature processing, model training, evaluation, registration, and promotion using tools such as Kubeflow, Airflow, MLflow, or equivalent
  • Automate model and service deployment with Docker, Kubernetes, Helm, and Terraform across AWS, Azure, or GCP environments
  • Develop CI/CD workflows for Python services, model artifacts, infrastructure changes, and container images using GitHub Actions, GitLab CI, or Jenkins
  • Implement production monitoring for service health, latency, throughput, resource utilization, data drift, model performance, and model quality regressions
  • Operate scalable model-serving infrastructure using platforms such as KServe, Seldon, NVIDIA Triton, Ray Serve, or managed cloud endpoints
  • Establish reproducibility and governance practices for datasets, features, model versions, experiments, secrets, and deployment approvals
  • Troubleshoot production incidents, improve system reliability, and document runbooks, architecture decisions, and operational standards
What We Are Looking For
  • 3–8 years of experience in MLOps, platform engineering, DevOps, or software engineering supporting machine learning systems in production
  • Strong Python and Linux skills, with experience building APIs, automation tools, and production services
  • Hands-on experience with Docker, Kubernetes, infrastructure as code, and at least one major cloud platform: AWS, Azure, or GCP
  • Practical knowledge of ML lifecycle tooling such as MLflow, Kubeflow, SageMaker, Vertex AI, Azure ML, Airflow, or comparable platforms
  • Experience designing CI/CD pipelines and observability using tools such as Prometheus, Grafana, OpenTelemetry, ELK, or cloud-native monitoring services
  • Bachelor’s degree in computer science, engineering, mathematics, or a related technical field; equivalent professional experience will be considered
  • Bonus: Experience with LLM serving, GPU scheduling, Ray, NVIDIA Triton, feature stores, GitOps, service meshes, model governance, or regulated production environments
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

MLOps Engineer
MLOps Engineer

Evlo AI • Seattle (WA)

On-site
USD 140,000 - 190,000
MLOps Engineer MLOps Engineer
MLOps Engineer MLOps Engineer

Kurai • Austin (TX)

On-site
USD 140,000 - 190,000
MLOps Engineer
MLOps Engineer

Evlo AI • Boston (MA)

On-site
USD 130,000 - 180,000
MLOps Engineer
MLOps Engineer

Sierracorp • San Francisco (CA)

On-site
USD 100,000 - 150,000
MLOps Engineer
MLOps Engineer

Compunnel, Inc. • San Antonio (TX)

On-site
USD 100,000 - 130,000
MLOps Engineer: Scalable ML Pipelines & Infra
MLOps Engineer: Scalable ML Pipelines & Infra

Compunnel, Inc. • San Antonio (TX)

On-site
MLOps Engineer
MLOps Engineer

Elevexa Career LLC • United States

On-site
USD 125,000 - 190,000
MLOps Engineer
MLOps Engineer

Arkhya Tech. Inc. • Scottsdale (AZ)

On-site
USD 140,000 - 180,000
Machine Learning Engineer
Machine Learning Engineer

AI Squared • Washington

On-site
USD 110,000 - 140,000
MLOps Engineer
MLOps Engineer

ACI Infotech • Atlanta (GA)

On-site
USD 100,000 - 120,000