AI Infrastructure Engineer - Hybrid, Impact & Innovation

Propio

Overland Park (KS)

Hybrid

USD 140,000 - 200,000

Full time

7 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Propio is seeking an AI Infrastructure Engineer in Overland Park, hybrid. You will design, build, and operate AWS-based, GPU-accelerated inference and streaming serving platforms for LLMs, ASR, and multimodal models, while supporting research training environments and ML/LLMOps capabilities.

You'll collaborate with researchers and edge teams to enable scalable, low-latency workloads and implement secure, observable infrastructure.

Qualifications

  • 3+ years building AI/ML infrastructure, inference platforms, or distributed systems in production.
  • Hands-on AWS experience with EKS, EC2 GPU workloads, ECR, S3, IAM/KMS, VPC networking, CloudWatch and/or OpenTelemetry.
  • Hands-on with at least one inference stack (vLLM, SGLang, TensorRT-LLM, Triton, KServe, or Ray Serve).
  • Experience operating ML/LLM systems in production, including model serving, autoscaling, monitoring, incident response, and performance benchmarking.
  • Familiarity with GPU infra and at least one serving stack (vLLM, SGLang, TensorRT-LLM, or Triton).
  • Working understanding of LLM Ops practices, including evaluation, observability and tracing, cost control, and versioning.

Responsibilities

  • Real-time inference systems: design, deploy, and operate low-latency streaming inference on AWS for LLMs and multimodal models.
  • Support research training environments: reproducible containers, GPU job scheduling, distributed execution, experiment tracking.
  • Build ML/LLMOps pipelines: registries, versioning, automated evaluation gates, deployments, canary and rollback workflows.
  • Develop Edge AI workflows: model optimization, packaging, validation, deployment for edge targets.
  • Ensure reliability and security: capacity planning, incident response, disaster recovery, IAM/KMS, private networking.

Skills

AI/ML infrastructure experience
AWS production experience
Inference stack knowledge
LLMOps practices understanding
GPU infrastructure familiarity
Production systems monitoring

Tools

vLLM
SGLang
TensorRT-LLM
Triton
KServe
Ray Serve
Megatron-LM
Slurm
SageMaker

Job description

Propio is seeking an AI Infrastructure Engineer in Overland Park, hybrid. You will design, build, and operate AWS-based, GPU-accelerated inference and streaming serving platforms for LLMs, ASR, and multimodal models, while supporting research training environments and ML/LLMOps capabilities.

You'll collaborate with researchers and edge teams to enable scalable, low-latency workloads and implement secure, observable infrastructure.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Infrastructure Engineer
AI Infrastructure Engineer

Propio • Overland Park (KS)

Hybrid
USD 140,000 - 200,000
Hybrid AI HPC Infrastructure Engineer (GPU/ML)
Hybrid AI HPC Infrastructure Engineer (GPU/ML)

Analysis Group, Inc. • Boston (MA)

On-site
USD 150,000 - 170,000
Discretionary annual bonus
Benefits package
Senior AI Engineer, Real-Time Production AI
Senior AI Engineer, Real-Time Production AI

Propio • Overland Park (KS)

Hybrid
USD 120,000 - 180,000
Senior AI Inference DevOps Engineer
Senior AI Inference DevOps Engineer

Lila Sciences • Cambridge (MA)

On-site
USD 192,000 - 272,000
Equity in company equity
Medical, dental, vision coverage
Generous PTO and holidays
Senior AI Compute Infra Engineer (Hybrid)
Senior AI Compute Infra Engineer (Hybrid)

Arm Limited • Seattle (WA)

Hybrid
USD 209,000 - 283,000
Relocation package
Recruitment accommodations
AI/ML Platform Engineer — Hybrid Cloud & GPU
AI/ML Platform Engineer — Hybrid Cloud & GPU

Madrona Venture Labs • United States

Hybrid
USD 180,000 - 260,000
AI Infrastructure Architect: Scalable GPU & LLM Deployments
AI Infrastructure Architect: Scalable GPU & LLM Deployments

Phase2 Technology • McLean (VA)

On-site
USD 112,800 - 257,000
Remote AI Infrastructure Engineer — Scale ML & Inference
Remote AI Infrastructure Engineer — Scale ML & Inference

Vantaca, LLC • Redwood City (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Medical, Dental, Vision
AI Infrastructure Engineer — Real-Time, Secure Cloud, Stock Options
AI Infrastructure Engineer — Real-Time, Secure Cloud, Stock Options

Palona AI • New York (NY)

On-site
USD 110,000 - 170,000
Competitive salary
Stock option plan
Medical, dental, vision benefits
+5
Senior AI Engineer
Senior AI Engineer

Propio • Overland Park (KS)

On-site
USD 120,000 - 180,000