GPUaaS Platform Engineer: Kubernetes & AI Ops

Veriipro

Irving (TX)

On-site

USD 140,000 - 180,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Veriipro in Irving, TX is seeking a GPUaaS Kubernetes Platform Engineer to design, operate, and support scalable GPU-enabled cloud infrastructure. You will manage Kubernetes and OpenShift clusters to enable AI/ML workloads, optimize GPU resource utilization, and ensure reliability of high-performance computing environments.

You will build and maintain CI/CD pipelines for GPU-based applications, automate platform operations, monitor utilization, and collaborate with AI/ML and DevOps teams to

Qualifications

  • Hands-on experience administering Kubernetes clusters and OpenShift environments.
  • Experience operating GPU-enabled Kubernetes platforms for AI/ML workloads.
  • Proficient in designing and implementing CI/CD pipelines for cloud platforms.
  • Strong knowledge of Linux system administration, networking, and storage.
  • Ability to create runbooks and operational documentation.

Responsibilities

  • Operate and maintain Kubernetes and OpenShift GPU platforms for AI/ML workloads.
  • Configure and manage GPU-enabled infrastructure, including GPU scheduling and resource optimization.
  • Develop and support GPUaaS platforms and scalable cloud infrastructure.
  • Implement CI/CD pipelines for GPU-based applications and infrastructure.
  • Ensure platform scalability, availability, and performance with monitoring and troubleshooting.
  • Collaborate with AI/ML and DevOps teams to deliver enterprise-grade GPU services.
  • Create dashboards, runbooks, and technical documentation.

Skills

Kubernetes administration
OpenShift
GPU workloads
CI/CD practices
Linux systems
Networking
Monitoring
Documentation

Education

Bachelor's degree in Computer Science or related field

Tools

Terraform
Ansible
GPU operators

Job description

Veriipro in Irving, TX is seeking a GPUaaS Kubernetes Platform Engineer to design, operate, and support scalable GPU-enabled cloud infrastructure. You will manage Kubernetes and OpenShift clusters to enable AI/ML workloads, optimize GPU resource utilization, and ensure reliability of high-performance computing environments.

You will build and maintain CI/CD pipelines for GPU-based applications, automate platform operations, monitor utilization, and collaborate with AI/ML and DevOps teams to

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPUaaS Kubernetes Platform Engineer
GPUaaS Kubernetes Platform Engineer

Veriipro • Irving (TX)

On-site
USD 140,000 - 180,000
GPU Platform Engineer for AI/ML Infra
GPU Platform Engineer for AI/ML Infra

Vero • United States

On-site
USD 136,000 - 160,000
Medical, dental, and vision insurance
Equity Scheme
401(k) with employer match
+3
Senior Kubernetes Engineer — GPU HPC & AI Platforms
Senior Kubernetes Engineer — GPU HPC & AI Platforms

NorthMark Compute & Cloud • Dallas (TX)

On-site
USD 140,000 - 210,000
Senior Platform Engineer - Kubernetes & GPU AI Infra
Senior Platform Engineer - Kubernetes & GPU AI Infra

2100 NVIDIA USA • Seattle (WA)

On-site
USD 272,000 - 431,000
Equity
Benefits
Remote AI Platform Engineer — Kubernetes & GPU, Equity
Remote AI Platform Engineer — Kubernetes & GPU, Equity

Hamilton Barnes Associates Limited • San Francisco (CA)

On-site
USD 250,000 - 300,000
Meaningful equity
Fully remote across North America
Full insurance coverage for you and你的依
AI Infra Engineer — GPU Cloud, Kubernetes/Slurm
AI Infra Engineer — GPU Cloud, Kubernetes/Slurm

Blue Signal Search • San Francisco (CA)

On-site
USD 180,000 - 240,000
Annual bonus
Equity participation
Comprehensive benefits
+1
Lead AI Infrastructure Engineer: Kubernetes & GPU
Lead AI Infrastructure Engineer: Kubernetes & GPU

Seekr • Washington

Hybrid
USD 180,000 - 230,000
Equity ownership
Unlimited PTO
Hybrid work (Reston, VA & Austin, TX)
GPU‑Optimized Kubernetes Engineer for AI/ML Platforms
GPU‑Optimized Kubernetes Engineer for AI/ML Platforms

NorthMark Compute and Cloud LLC • Dallas (TX)

On-site
USD 120,000 - 160,000
Company-Paid Lunch Stipend via GrubHub
100% Employer-Paid Medical, Dental, and Vision
16 weeks Paid Parental Leave
+2
AI Infrastructure Engineer — GPU Kubernetes for Production
AI Infrastructure Engineer — GPU Kubernetes for Production

vCluster • Germany (OH)

On-site
USD 150,000 - 200,000
Competitive Salary
Platinum-Level Insurance
Flexible Working Schedule
+1
Senior AI Infra Engineer: GPU Compute on Kubernetes
Senior AI Infra Engineer: GPU Compute on Kubernetes

Harell Data • Palo Alto (CA)

On-site
USD 180,000 - 260,000