AI Operations Platform Consultant

MACHINE LEARNING TECHNOLOGIES LLC

Jersey City (NJ)

On-site

USD 70,000 - 100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

MACHINE LEARNING TECHNOLOGIES LLC is seeking an AI Operations Platform Consultant experienced in deploying and managing large-scale GPU-accelerated AI platforms. Responsibilities include leading LLMOps processes and optimizing LLM pipelines for performance across Kubernetes environments.

The ideal candidate must possess strong expertise in Triton Inference Server and TensorRT-LLM, and will be responsible for building production-grade LLM inference systems. This position offers competitive hourly rates and the opportunity for extension beyond 24 months.

Qualifications

  • Extensive experience in deploying and managing LLM inference systems.
  • Strong expertise in Triton Inference Server and TensorRT-LLM.
  • Experience in optimizing production-grade LLM pipelines.

Responsibilities

  • Lead end-to-end LLMOps processes with model versioning and automated rollouts.
  • Manage AI inference service monitoring for performance.
  • Optimize LLM models using techniques like quantization and pruning.

Skills

Large-scale GPU-accelerated AI platforms
LLM inference systems on Kubernetes
Triton Inference Server
TensorRT-LLM
MLOps/LLMOps pipelines
Containerized services

Job description

Overview

Job Description

Job Description:

  • Job ID: J53022
  • Job Title: AI Operations Platform Consultant
  • Location: Jersey City, NJ
  • Duration: 24 Months + Extension
  • Hourly Rate: Depending on Experience (DOE)
  • Work Authorization: US Citizen, Green Card, OPT-EAD, CPT, H-1B, H4-EAD, L2-EAD, GC-EAD
  • Client: To Be Discussed Later
  • Employment Type: W-2, 1099, C2C
Job Description
  • Brings extensive experience operating large-scale GPU-accelerated AI platforms, deploying and managing LLM inference systems on Kubernetes with strong expertise in Triton Inference Server and TensorRT-LLM.
  • Experience building and optimizing production-grade LLM pipelines with GPU-aware scheduling, load balancing, and real-time performance tuning across multi-node clusters. Design of containerized microservices, deployment workflows, and maintaining operational reliability in mission-critical environments.
  • Led end-to-end LLMOps processes involving model versioning, engine builds, automated rollouts, and secure runtime controls.
  • Developed comprehensive observability for inference systems, using telemetry and dashboards to track GPU health, latency, throughput, and service availability.
  • Applied advanced optimization methods such as mixed precision, quantization, sharding, and batching to improve efficiency. Strong blend of platform engineering, AI infrastructure, and hands-on operational experience running high-performance LLM systems in production.
Basic Info
  • AI Operations Platform Consultant
  • Experience deploying, managing, operating, and troubleshooting containerized services at scale on Kubernetes for mission-critical applications (OpenShift)
  • Experience deploying, configuring, and tuning LLMs using TensorRT-LLM and Triton Inference Server
  • Managing MLOps/LLMOps pipelines, deploying inference services in production
  • Setup and operation of AI inference service monitoring for performance and availability
  • Experience deploying and troubleshooting LLM models on a containerized platform, monitoring, load balancing
  • Operation and support of MLOps/LLMOps pipelines for production
  • Experience with incident management, change management, event management for mission-critical systems
  • Managing scalable infrastructure for deploying and managing LLMs
  • Deploying models in production environments, including containerization, microservices, and API design
  • Triton Inference Server architecture, configuration, and deployment
  • Model optimization techniques using Triton with TRTLLM
  • Model optimization techniques including pruning, quantization, and knowledge distillation
Equal Opportunity

MACHINE LEARNING TECHNOLOGIES LLC is an equal opportunity employer inclusive of female, minority, disability and veterans, (M/F/D/V). Hiring, promotion, transfer, compensation, benefits, discipline, termination and all other employment decisions are made without regard to race, color, religion, sex, sexual orientation, gender identity, age, disability, national origin, citizenship/immigration status, veteran status or any other protected status. MACHINE LEARNING TECHNOLOGIES LLC will not make any posting or employment decision that does not comply with applicable laws relating to labor and employment, equal opportunity, employment eligibility requirements or related matters. Nor will MACHINE LEARNING TECHNO Technologies LLC require in a posting or otherwise U.S. citizenship or lawful permanent residency in the U.S. as a condition of employment except as necessary to comply with law, regulation, executive order, or federal, state, or local government contract.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Operations Platform Consultant
AI Operations Platform Consultant

Cloud Analytics Technologies, LLC • Jersey City (NJ)

On-site
USD 130,000 - 160,000
AI Operations Platform Engineer - LLMs on Kubernetes
AI Operations Platform Engineer - LLMs on Kubernetes

MACHINE LEARNING TECHNOLOGIES LLC • Jersey City (NJ)

On-site
USD 70,000 - 100,000
On-Prem LLM Platform Engineer (OpenShift AI / GPU)
On-Prem LLM Platform Engineer (OpenShift AI / GPU)

Infosys Limited • Charlotte (NC)

On-site
USD 80,000 - 120,000
Long-term disability
Health reimbursement accounts
Insurance offerings
+1
LLMOps Platform Engineer for GPU AI Inference
LLMOps Platform Engineer for GPU AI Inference

Cloud Analytics Technologies, LLC • Jersey City (NJ)

On-site
USD 130,000 - 160,000
Tech Lead Software Engineer - AI Compute Infrastructure
Tech Lead Software Engineer - AI Compute Infrastructure

ByteDance • Seattle (WA)

On-site
USD 232,560 - 427,500
AI / ML Engineer
AI / ML Engineer

Machine Intelligence Technologies, LLC • San Jose (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Equal opportunity employer
ML engineer with Gen AI
ML engineer with Gen AI

MACHINE LEARNING TECHNOLOGIES LLC • Austin (TX)

On-site
USD 83,000 - 165,000
Senior Developer
Senior Developer

ICE Clear Europe Limited • Atlanta (GA)

On-site
USD 150,000 - 210,000
Senior AI Developer
Senior AI Developer

MACHINE LEARNING TECHNOLOGIES LLC • Mettawa (IL)

On-site
USD 110,000 - 150,000
Equal Opportunity Employer
Inclusive workplace for minorities and veterans
Senior AI Cloud Architect
Senior AI Cloud Architect

MACHINE LEARNING TECHNOLOGIES LLC • Minneapolis (MN)

On-site
USD 120,000 - 160,000