Senior MLOps Engineer: Scalable AI Deployments

Evlo AI

Atlanta (GA)

On-site

USD 120,000 - 190,000

Full time

19 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Evlo AI is seeking an experienced MLOps/DevOps engineer to own the infrastructure and CI/CD pipelines powering its large-scale ML systems in Atlanta, GA. You will ensure smooth transitions from research notebooks to production services and collaborate with data scientists, ML engineers, and platform architects to build scalable, automated deployment workflows.

You will monitor model performance and data drift, optimize serving with Triton or TorchServe, and enforce security and governance for

Qualifications

  • 3–6 years of experience in MLOps, DevOps, or reliability engineering with a heavy focus on machine learning workloads.
  • Strong proficiency in Python, Bash, and infrastructure-as-code tools such as Terraform or CloudFormation
  • Deep hands-on experience with containerization and orchestration platforms, particularly Docker and Kubernetes
  • Familiarity with feature stores and model registries like Feast, Hopsworks, or MLflow
  • Solid understanding of CI/CD pipelines for software and ML codebases (GitHub Actions, GitLab CI, ArgoCD)
  • Bonus: Experience managing LLM inference pipelines, vLLM, TensorRT-LLM, or large-scale distributed training clusters

Responsibilities

  • Design, build, and maintain robust MLOps infrastructure using Kubernetes, Docker, Terraform, and cloud-native tools
  • Automate end-to-end model training, validation, and deployment pipelines utilizing tools like MLflow, Kubeflow, or AWS SageMaker
  • Implement comprehensive monitoring systems to track model performance, latency, throughput, data drift, and concept drift in real time
  • Optimize model serving architectures for low latency and high concurrency using Triton Inference Server or TorchServe
  • Establish security, governance, and access control best practices for data storage, feature stores, and model artifact registries
  • Collaborate with engineering teams to troubleshoot production incidents and continuously improve system reliability

Skills

Python
Bash
CI/CD pipelines

Tools

Kubernetes
Docker
Terraform
CloudFormation
MLflow
Kubeflow
AWS SageMaker
Feast
Hopsworks
TorchServe
Triton Inference Server
vLLM
TensorRT-LLM
ArgoCD
GitHub Actions
GitLab CI

Job description

Evlo AI is seeking an experienced MLOps/DevOps engineer to own the infrastructure and CI/CD pipelines powering its large-scale ML systems in Atlanta, GA. You will ensure smooth transitions from research notebooks to production services and collaborate with data scientists, ML engineers, and platform architects to build scalable, automated deployment workflows.

You will monitor model performance and data drift, optimize serving with Triton or TorchServe, and enforce security and governance for

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior MLOps Engineer: Scalable AI Production & DevOps
Senior MLOps Engineer: Scalable AI Production & DevOps

AgileEngine • Dallas (TX)

On-site
USD 120,000 - 180,000
Professional growth
Competitive USD-based compensation
Exciting projects
+1
Remote MLOps Engineer: Scale AI Deployments & Pipelines
Remote MLOps Engineer: Scale AI Deployments & Pipelines

AgileEngine, LLC. • Jacksonville (TX)

On-site
USD 120,000 - 180,000
Growth without limits
Competitive compensation
Flexibility
+3
Remote MLOps Engineer: Scale ML Deployments and CI/CD
Remote MLOps Engineer: Scale ML Deployments and CI/CD

AgileEngine • Carrollton (TX)

On-site
USD 140,000 - 190,000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3
Production ML Engineer: Build, Deploy & Monitor AI
Production ML Engineer: Build, Deploy & Monitor AI

Evlo AI • Raleigh (NC)

On-site
USD 120,000 - 180,000
Remote MLOps Engineer | Scale AI/ML in Production
Remote MLOps Engineer | Scale AI/ML in Production

AgileEngine • Dallas (TX)

On-site
USD 120,000 - 180,000
Growth opportunities
Competitive compensation
Remote work
+3
Remote MLOps Engineer - Scale AI Deployments
Remote MLOps Engineer - Scale AI Deployments

AgileEngine, LLC. • Baltimore (MD)

On-site
USD 140,000 - 190,000
Growth without limits
Competitive compensation
Flexibility: 100% remote with flexible
+3
Remote MLOps Engineer: Scale Production AI Pipelines
Remote MLOps Engineer: Scale Production AI Pipelines

AgileEngine, LLC. • Richmond (VA)

On-site
USD 120,000 - 180,000
Growth without limits
Competitive compensation
Flexibility — 100% remote with hours
+3
Remote MLOps Engineer: Scale Production ML Pipelines
Remote MLOps Engineer: Scale Production ML Pipelines

AgileEngine, LLC. • West Palm Beach (FL)

On-site
USD 120,000 - 180,000
Growth opportunities
Competitive compensation
Remote work
+3
Production ML Engineer: Scalable Pipelines and Low-Latency
Production ML Engineer: Scalable Pipelines and Low-Latency

Evlo AI • Austin (TX)

On-site
USD 120,000 - 180,000
MLOps Engineer: Deploy, Observe & Scale AI
MLOps Engineer: Deploy, Observe & Scale AI

Speria • Atlanta (GA)

On-site
USD 120,000 - 180,000