MLOps Engineer

InfoVision Inc.

Irving (TX)

On-site

USD 100,000 - 130,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

InfoVision Inc. is seeking an MLOps Engineer to productionize and scale machine learning and Generative AI systems. The successful candidate will focus on LLM deployment, orchestration, and reliability in production environments.

This role involves deploying and managing ML/DL models, building Kubernetes-based infrastructures, and designing scalable inference systems. Strong experience with model deployment, optimization of LLMs, and proficiency in Python are essential.

Qualifications

  • Experience with deploying, managing, and scaling ML/DL models in production.
  • Hands-on experience with Kubernetes-based infrastructure for ML workloads.
  • Proficiency in model packaging, serialization, and versioning.

Responsibilities

  • Deploy, manage, and scale ML/DL models in production.
  • Build and operate Kubernetes-based infrastructure for ML workloads.
  • Design scalable inference systems for batch and real-time processing.

Skills

Hands-on experience with ML/DL models and serialization
Proven experience in model deployment, scaling, and monitoring
Experience with local LLM deployment and optimization
Solid understanding of LLM memory patterns
Experience with API gateways
Familiarity with GenAI workflows
Experience building agentic systems
Proficiency in Python

Job description

MLOps Engineer to productionize and scale ML and GenAI systems, with a focus on LLM deployment, orchestration, and reliability in production environments.

Key Responsibilities

Deploy, manage, and scale ML/DL models in production

Build and operate Kubernetes-based infrastructure for ML workloads

Handle model packaging, serialization, and versioning

Design scalable inference systems (batch and real-time)

Deploy and optimize local LLMs (latency, throughput, cost)

Build and manage agentic systems with tool integration

Design and manage LLM memory (short-term, long-term, vector stores)

Integrate and manage API gateways for model access, routing, and rate limiting

Monitor performance, drift, and system reliability

Requirements

Hands-on experience with ML/DL models and serialization

Proven experience in model deployment, scaling, and monitoring

Experience with local LLM deployment and optimization

Solid understanding of LLM memory patterns (context windows, retrieval, persistence)

Experience with API gateways, load balancing, and service routing

Familiarity with GenAI workflows (RAG, orchestration frameworks)

Experience building agentic / multi-step LLM systems

Proficiency in Python and modern ML/infra tooling

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

MLOps Engineer
MLOps Engineer

Sierracorp • San Francisco (CA)

On-site
USD 100,000 - 150,000
MLOps Engineer MLOps Engineer
MLOps Engineer MLOps Engineer

Kurai • Austin (TX)

On-site
USD 140,000 - 190,000
AI / ML Ops Engineer
AI / ML Ops Engineer

Zoho • United States

Remote
USD 140,000 - 210,000
MLOps Engineer
MLOps Engineer

Evlo AI • Seattle (WA)

On-site
USD 130,000 - 190,000
MLOps Engineer: GenAI & LLM Production
MLOps Engineer: GenAI & LLM Production

InfoVision Inc. • Irving (TX)

On-site
USD 100,000 - 130,000
MLOps Engineer
MLOps Engineer

Soledad Data Solutions • California (MO)

On-site
USD 140,000 - 210,000
MLOps Engineer: Scalable ML Pipelines & Infra
MLOps Engineer: Scalable ML Pipelines & Infra

Compunnel, Inc. • San Antonio (TX)

On-site
Confidential
MLOps Engineer
MLOps Engineer

Compunnel, Inc. • San Antonio (TX)

On-site
USD 100,000 - 130,000
MLOps Engineer
MLOps Engineer

ACI Infotech • Atlanta (GA)

On-site
USD 100,000 - 120,000
MLOPs Architect
MLOPs Architect

Quantum World Technologies Inc. • Dallas (TX)

On-site
USD 120,000 - 180,000