Machine Learning Operations (MLOps) Engineer

Openlane

Carmel, Northern (IN, KY)

Híbrido

USD 90.000 - 130.000

Jornada completa

Hace 2 días
Sé de los primeros/as/es en solicitar esta vacante
Generador de candidaturas

Destaca para este puesto — genera un currículum y una carta de presentación adaptados en cuestión de un minuto.

Supera los filtros ATS

Ventajas ofrecidas por este puesto de trabajo

Medical benefits
401K matching
Paid time off

Descripción de la vacante

OPENLANE is seeking an experienced MLOps Engineer to build and operate scalable ML pipelines from experimentation to production. You will own automated workflows, model deployment, monitoring, and safe updates while reducing toil through reusable platform components.

You will collaborate with data scientists, engineers, and SRE to ensure reliable, cost-efficient ML delivery across cloud environments. A strong Python background and cloud experience are essential.

Formación

  • 3+ years in MLOps, ML engineering, or similar, with production ML pipelines.
  • Bachelor’s degree or equivalent practical experience.
  • Strong Python and cloud experience (AWS preferred).
  • Experience with workflow orchestration, IaC, CI/CD, containers, and monitoring.

Responsabilidades

  • Design, build, and operate automated ML pipelines for data ingestion, training, validation, deployment, and rollback.
  • Implement model CI/CD with testing, canary, and safe rollout strategies.
  • Monitor model performance, data drift, latency, and availability with actionable alerts.
  • Ensure reproducibility of workloads via versioned code, data, features, and environments.

Conocimientos

Python
Cloud platforms
CI/CD for ML
Observability
Automation
AI tooling evaluation

Educación

Bachelor’s degree

Herramientas

Airflow
Kubeflow
Terraform
Docker
Kubernetes
TorchServe / TF Serving

Descripción del empleo

## Machine Learning Operations (MLOps) EngineerApply: Remote: US - IN - Carmel (OPENLANE): Full time: Posted Today: R-255077**Who We Are:** At OPENLANE we make wholesale easy so our customers can be more successful. **We’re a technology company** building the world’s most advanced—and uncomplicated—digital marketplace for used vehicles. **We’re a data company** helping customers buy and sell smarter with clear, actionable insights they can understand and use. **And we’re an innovation company** accelerating the future of wholesale remarketing through curiosity, collaboration, and an entrepreneurial spirit. **Our Values:** **Driven Waybuilders.** We pursue challenges that inspire us to build, create, and innovate. **Relentless Curiosity.** We seek to understand and improve our customers’ experience. **Smart Risk-Taking.** We transform risk into progress through data, experience, and intuition. **Fearless Ownership.** We deliver what we promise and learn along the way.**We’re Looking For**An Machine Learning Operations (MLOps) Engineer who builds and operates the pipelines, platforms, and observability that carry machine learning models from experimentation into reliable, cost-efficient production. You will join a new MLOps team within Site Reliability Engineering and help define how OPENLANE trains, ships, monitors, and safely updates models at scale. You will independently design automated ML delivery workflows, reduce operational toil, and ensure models are reproducible, monitored, and serving predictions safely in production. **You Are**Analytical, pragmatic, proactive, collaborative, quality-focused, and AI-forward. You treat AI as a first-class engineering tool, critically evaluate its output, and apply strong engineering judgment to systems that must stay healthy long after the first deployment. **You Bring*** Core knowledge of system architecture, distributed systems, cloud-native design, networking, and agile development methodologies (Scrum, Kanban).* Functional expertise in designing end-to-end ML pipelines (data ingestion, training, validation, deployment), including requirements for latency, redundancy, scalability, and error handling.* Technical expertise in CI/CD for ML, infrastructure-as-code, containerization and orchestration, and metrics, monitoring, and alerting toolsets.* Working understanding of the model lifecycle: experiment tracking and lineage, model registries, model drift, data quality, and retraining triggers, with enough ML fluency to diagnose model issues alongside data scientists.* Strong Python skills and the ability to write clean, well-tested, production-ready code. **You Own*** Design, build, and operation of automated ML pipelines covering data ingestion, training, validation, deployment, and rollback.* Model CI/CD, including automated testing and canary and rollback strategies for safe releases.* Monitoring and observability for model performance, data drift, inference latency, and availability, with actionable alerting and incident response.* Reproducibility of model workloads through versioned code, data, features, and environments.* Identification and reduction of operational toil through automation, standardization, and reusable ML platform components.* Documentation, testing, and debugging of pipelines and serving infrastructure, maintaining high code and design standards. **You Drive****Execution:** Translate data science and product needs into scalable, automated ML delivery designs, and implement technical roadmap projects for the MLOps platform.**Results:** Shorten time-to-production for models while improving reliability, minimizing failure risk, and optimizing the cost, scaling, and utilization of training and inference workloads in partnership with Infrastructure teams.**Team:** Support hiring, knowledge sharing, and mentorship as the MLOps practice grows. **How You’ll Use AI in This Role*** AI-assisted coding, refactoring, and test generation for pipelines and platform tooling.* AI-supported debugging and root-cause analysis of pipeline failures and model degradation.* AI-assisted documentation and runbooks to reduce repetitive work and increase engineering leverage. **Who You Will Work With**Reporting to the Director, Site Reliability Engineering, this role collaborates daily with Data Scientists, Data Engineers, Software Engineers, Infrastructure, Security, Product Owners, and SRE teams across the organization. **Where You Work**Your work is performed as a Remote employee, with occasional team alignment meetings as needed. **Must Have’s*** 3+ years of experience in MLOps, ML engineering, Site Reliability Engineering, DevOps, or Software Engineering, including experience building or operating ML pipelines in production.* Bachelor’s degree preferred or equivalent practical experience.* Strong Python skills and hands-on experience with a major cloud platform (AWS preferred).* Experience with workflow orchestration (e.g., Airflow, Kubeflow, Step Functions), infrastructure-as-code (e.g., Terraform), and CI/CD systems.* Experience with containers and Kubernetes (Docker, EKS preferred).* Familiarity with model serving frameworks (e.g., TorchServe, TF Serving, Triton, BentoML, SageMaker endpoints) and monitoring tools (e.g., Prometheus, Grafana, Evidently AI).* Hands-on experience with system debugging, observability, and incident response in production environments.* Actively uses AI development tools and can critically evaluate AI outputs. **Nice to Have’s*** Experience with ML platforms and tooling such as SageMaker, MLflow, Databricks, or Vertex AI (GCP).* Experience with feature stores, model registries, or LLM/GenAI deployment and evaluation (LLMOps).* Experience with GPU workload scheduling and inference cost optimization.* Experience interviewing candidates, conducting technical debriefs, and mentoring junior engineers or interns.* Proven ability to acquire deep domain and architectural knowledge within 6 months of joining a team.**What We Offer:*** Competitive pay* Medical, dental, and vision benefits with employer HSA contributions (US) and FSA options (US)* Immediately vested 401K (US) or RRSP (Canada) with company match* Paid Vacation, Personal, and Sick Time* Paid maternity and paternity leave (US)* Employer-paid short-term disability, long-term disability, life insurance, and AD&D (US)* Robust Employee Assistance Program* Employer paid Leap into Service Day to volunteer* Tuition Reimbursement for eligible programs* Opportunities to expand your skill set and share your knowledge across a publicly traded, global organization* Company culture of internal promotions, diverse career paths, and meaningful advancement**Sound like a match? Apply Now - We can't wait to hear from you!**
Consigue la evaluación confidencial y gratuita de tu currículum.

o arrastra y suelta tu archivo aquí

Similar jobs

Puestos de trabajo similares que vale la pena comparar

AI First Software Engineer
AI First Software Engineer

Openlane • Carmel (IN)

Híbrido
USD 120.000 - 180.000
Medical, dental, and vision benefits
Employer HSA contributions
Immediately vested 401K with company 1
+4
Machine Learning Operations (MLOps) Engineer
Machine Learning Operations (MLOps) Engineer

OPEN OPENLANE US Inc • Carmel (IN)

A distancia
USD 110.000 - 165.000
Medical, dental, and vision benefits
401K with company match
Paid vacation and sick time
+4
Machine Learning Operations Engineer
Machine Learning Operations Engineer

Imagineteam • Charlotte (NC), Northern (KY)

Híbrido
USD 120.000 - 180.000
AI/ML Engineer
AI/ML Engineer

GemPixel Inc. • EE. UU.

A distancia
USD 110.000 - 170.000
Remote-friendly
International team
Technology budget
Senior AI Engineer - USA
Senior AI Engineer - USA

Cogniify, Inc. • San Francisco (CA)

Presencial
USD 140.000 - 165.000
Unlimited PTO
Generous parental leave
Entrepreneurial culture
+6
Software Engineer - USA
Software Engineer - USA

Cogniify, Inc. • Santa Clara (CA), Northern (KY)

Híbrido
USD 150.000 - 170.000
Unlimited PTO
Parental leave (generous)
Medical insurance
+3
Architect - Platform Engineering - USA
Architect - Platform Engineering - USA

Quantiphi, Inc. • Chicago (IL)

Presencial
USD 130.000 - 180.000
Opportunity to work with Fortune 500 companies
Exposure to cutting-edge AI technologies
Dynamic team environment
Senior ML OPs Engineer
Senior ML OPs Engineer

Glocomms • California (MO)

Presencial
USD 198.000 - 230.000
Meal stipends for remote work days
Generous paid time off
Comprehensive health coverage
+1
ML Ops Engineer — Agentic AI Lab (Founding Team)
ML Ops Engineer — Agentic AI Lab (Founding Team)

Fabrion • San Francisco (CA)

Presencial
USD 120.000 - 150.000
Competitive salary
Meaningful equity
Senior Machine Learning Engineer
Senior Machine Learning Engineer

Regenerativeaitool • Boston (MA)

Presencial
USD 120.000 - 150.000
Comprehensive health, dental, and vision insurance
Flexible PTO and remote-friendly work arrangements
Annual learning and development budget ($5,000)
+2