LLMOps Platform Engineer for GPU AI Inference

Cloud Analytics Technologies, LLC

Jersey City (NJ)

On-site

USD 130,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading AI infrastructure company located in New Jersey is seeking an experienced AI Operations Platform Consultant to lead and optimize large-scale GPU-accelerated AI platforms. The ideal candidate will have a strong background in deploying and managing LLM inference systems on Kubernetes, with expertise in TensorRT-LLM and Triton Inference Server. Responsibilities include managing production-grade LLM pipelines and ensuring operational reliability. This position is part of a team committed to diversity and equal opportunity.

Qualifications

  • Extensive experience operating large-scale GPU-accelerated AI platforms.
  • Strong expertise in deploying and managing LLM inference systems.
  • Experience with AI inference service monitoring and performance optimization.

Responsibilities

  • Lead production-grade LLM pipelines with performance tuning.
  • Manage MLOps processes for deploying inference services.
  • Develop observability for inference systems using telemetry.

Skills

Kubernetes
TensorRT-LLM
Triton Inference Server
MLOps
AI operations
Containerization

Job description

A leading AI infrastructure company located in New Jersey is seeking an experienced AI Operations Platform Consultant to lead and optimize large-scale GPU-accelerated AI platforms. The ideal candidate will have a strong background in deploying and managing LLM inference systems on Kubernetes, with expertise in TensorRT-LLM and Triton Inference Server. Responsibilities include managing production-grade LLM pipelines and ensuring operational reliability. This position is part of a team committed to diversity and equal opportunity.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI Operations Platform Consultant
AI Operations Platform Consultant

Cloud Analytics Technologies, LLC • Jersey City (NJ)

On-site
USD 130,000 - 160,000
AI Infrastructure Engineer - Inference Platform
AI Infrastructure Engineer - Inference Platform

Hoonify Technologies Inc. • Albuquerque (NM)

On-site
USD 120,000 - 190,000
Staff Engineer, LLM Inference & Infra
Staff Engineer, LLM Inference & Infra

Prime Intellect • United States

Hybrid
USD 120,000 - 150,000
Competitive compensation
Flexible work arrangement
Full visa sponsorship
+2
ML Engineer—LLM Inference & GPU Optimization (Equity)
ML Engineer—LLM Inference & GPU Optimization (Equity)

IC Resources • San Francisco (CA)

On-site
USD 200,000 - 290,000
401(k)
Unlimited PTO
Modern engineering workspace
+1
On-Prem LLM Inference Engineer: GPU & AI Infra
On-Prem LLM Inference Engineer: GPU & AI Infra

Compunnel, Inc. • Charlotte (NC)

On-site
USD 120,000 - 150,000
LLM Inference GPU Systems Consultant
LLM Inference GPU Systems Consultant

Delan Associates, Inc • Charlotte (NC)

On-site
USD 150,000 - 210,000
On-Premise LLM Inference & GPU Systems Engineer
On-Premise LLM Inference & GPU Systems Engineer

NTT DATA North America • Charlotte (NC)

On-site
USD 120,000 - 150,000
AI Inference Platform Engineer
AI Inference Platform Engineer

Hoonify Technologies Inc. • Albuquerque (NM)

On-site
USD 120,000 - 190,000
Onsite LLM Inference Architect for NVIDIA GPU Infra
Onsite LLM Inference Architect for NVIDIA GPU Infra

Delan Associates, Inc • Charlotte (NC)

On-site
USD 150,000 - 210,000
Remote MLOps Engineer - Scalable AI Inference Platform
Remote MLOps Engineer - Scalable AI Inference Platform

Bright Vision Technologies • Sammamish (WA)

On-site
USD 100,000 - 150,000