MLOps Engineer

ERT, Inc.

Arlington (VA)

On-site

USD 120,000 - 150,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ERT, Inc. is seeking an MLOps Engineer in Arlington, Virginia. This role involves ensuring the seamless deployment, optimization, and monitoring of AI models in production. Key responsibilities include managing machine learning models, building scalable infrastructures, and implementing monitoring tools. The ideal candidate will have a Bachelor's or Master's degree in a relevant field, along with 5+ years of experience in MLOps. Candidates must be eligible for Department of Homeland Security clearance. ERT, Inc. is an equal opportunity employer.

Qualifications

  • 5+ years in MLOps, DevOps, or software engineering with a focus on AI/ML systems.
  • Proven experience deploying models in production using MLflow, Kubeflow, or cloud platforms.
  • Hands-on experience with observability tools for real-time monitoring.
  • Strong problem-solving and debugging skills for resolving pipeline and monitoring issues.
  • Must be eligible to obtain a Department of Homeland Security EOD clearance.

Responsibilities

  • Deploy and manage machine learning models in production ensuring scalability.
  • Build and maintain dashboards to track model health and historical trends.
  • Implement drift detection pipelines to identify shifts in data distributions.
  • Set up centralized logging to capture AI inference events and audit trails.
  • Develop CI/CD pipelines to automate model updates and deployment.

Skills

Python
SQL
Containerization (Docker, Kubernetes)
CI/CD tools (GitHub Actions, Jenkins)
Machine Learning
Observability tools (Prometheus, Grafana)
Drift detection tools (Evidently AI, Alibi Detect)
Time-series databases (InfluxDB, TimescaleDB)
Visualization libraries (Plotly, Seaborn)
JavaScript or Go

Education

Bachelor's or Master's degree in Computer Science, Data Science, Engineering, or a related field

Tools

MLflow
Kubeflow
AWS SageMaker
Kibana
ELK Stack
OpenTelemetry

Job description

Overview

We are seeking a skilled MLOps Engineer to join our team and ensure the seamless deployment, monitoring, and optimization of AI models in production. The MLOps Engineer will design, implement, and maintain end‑to‑end machine learning pipelines, focusing on automating model deployment, monitoring model health, detecting data drift, and managing AI‑related logging. This role will involve building scalable infrastructure and dashboards for real‑time and historical insights, ensuring models are secure, performant, and aligned with business needs.

Key Responsibilities
  • Deploy and manage machine learning models in production using tools like MLflow, Kubeflow, or AWS SageMaker, ensuring scalability and low latency.
  • Build and maintain dashboards using Grafana, Prometheus, or Kibana to track real‑time model health (accuracy, latency) and historical trends.
  • Implement drift detection pipelines using tools like Evidently AI or Alibi Detect to identify shifts in data distributions and trigger alerts or retraining.
  • Set up centralized logging with ELK Stack or OpenTelemetry to capture AI inference events, errors, and audit trails for debugging and compliance.
  • Develop CI/CD pipelines with GitHub Actions or Jenkins to automate model updates, testing, and deployment.
  • Apply secure‑by‑design principles to protect data pipelines and models, using encryption, access controls, and compliance with regulations like GDPR or NIST AI RMF.
  • Collaborate with data scientists, AI Integration Engineers, and DevOps teams to align model performance with business requirements and infrastructure capabilities.
  • Optimize models for production (e.g., via quantization or pruning) and ensure efficient resource usage on cloud platforms like AWS, Azure, or Google Cloud.
  • Maintain clear documentation of pipelines, dashboards, and monitoring processes for cross‑team transparency.
Qualifications
  • Education: Bachelor's or Master's degree in Computer Science, Data Science, Engineering, or a related field.
  • Experience:
    • 5+ years in MLOps, DevOps, or software engineering with a focus on AI/ML systems.
    • Proven experience deploying models in production using MLflow, Kubeflow, or cloud platforms (AWS SageMaker, Azure ML).
    • Hands‑on experience with observability tools like Prometheus, Grafana, or Datadog for real‑time monitoring.
  • Technical Skills:
    • Proficiency in Python and SQL; familiarity with JavaScript or Go is a plus.
    • Expertise in containerization (Docker, Kubernetes) and CI/CD tools (GitHub Actions, Jenkins).
    • Knowledge of time‑series databases (InfluxDB, TimescaleDB) and logging frameworks (ELK Stack, OpenTelemetry).
    • Experience with drift detection tools (Evidently AI, Alibi Detect) and visualization libraries (Plotly, Seaborn).
  • AI‑Specific Skills:
    • Understanding of model performance metrics (precision, recall, AUC) and drift detection methods (KS test, PSI).
    • Familiarity with AI vulnerabilities (data poisoning, adversarial attacks) and mitigation tools like Adversarial Robustness Toolbox (ART).
  • Soft Skills:
    • Strong problem‑solving and debugging skills for resolving pipeline and monitoring issues.
    • Excellent collaboration and communication skills to work with cross‑functional teams.
    • Attention to detail for ensuring accurate and secure dashboard reporting.
  • Must be eligible to obtain a Department of Homeland Security EOD clearance (Requirements: 1. US Citizenship, 2. Favorable Background Investigation).
Preferred Qualifications
  • Experience with LLM monitoring tools like LangSmith or Helicone for generative AI applications.
  • Knowledge of compliance frameworks (GDPR, HIPAA) for secure data handling.
  • Contributions to open‑source MLOps projects or familiarity with X platform discussions on MLOps or AIOps.
Equal Opportunity Statement

Entarian is an Equal Opportunity and Affidavitative Action Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, pregnancy, sexual orientation, gender identity, national origin, age, protected veteran status, or disability status.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

MLOps Engineer
MLOps Engineer

Sierracorp • San Francisco (CA)

On-site
USD 100,000 - 150,000
ML Operations Engineer
ML Operations Engineer

NextGen Healthcare • Georgia

On-site
USD 80,000 - 120,000
MLOps Engineer
MLOps Engineer

Compunnel, Inc. • San Antonio (TX)

On-site
USD 100,000 - 130,000
MLOps Engineer
MLOps Engineer

Codinix Consulting Services • California (MO)

On-site
USD 120,000 - 150,000
MLOps Engineer: Scalable ML Pipelines & Infra
MLOps Engineer: Scalable ML Pipelines & Infra

Compunnel, Inc. • San Antonio (TX)

On-site
MLOps Engineer
MLOps Engineer

XM • Town of Poland (NY)

On-site
USD 120,000 - 180,000
Private health insurance
International training opportunities
MLOps Engineer
MLOps Engineer

Blue Signal Search • Santa Clara (CA)

On-site
USD 140,000 - 190,000
Advanced GPU infra exposure
Collaborative engineering culture
Open source AI frameworks access
+2
MLOps Engineer MLOps Engineer
MLOps Engineer MLOps Engineer

Kurai • Austin (TX)

On-site
USD 140,000 - 190,000
MLOps Engineer - Scalable ML Pipelines & CI/CD
MLOps Engineer - Scalable ML Pipelines & CI/CD

Codinix Consulting Services • California (MO)

On-site
MLOPs Architect
MLOPs Architect

Quantum World Technologies Inc. • Dallas (TX)

On-site
USD 120,000 - 180,000