AI Platform Engineer: Scale GPU AI & Research

Johns Hopkins Applied Physics Lab

Laurel (MD)

On-site

USD 85,000 - 165,000

Full time

8 days ago
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Johns Hopkins Applied Physics Laboratory seeks an AI Solutions Engineer to bridge researchers and a scalable AI computing platform. You will configure and evolve the AI stack, collaborate on platform deployments, and assist researchers with GPU-accelerated workloads for training, inference, and deployment.

Ideal candidates have 3+ years in Kubernetes/Docker, hands-on Jupyter, and experience with PyTorch, TensorFlow, or similar frameworks.

Qualifications

  • Bachelor of Science degree or equivalent professional experience.
  • 3+ years of container-based orchestration (Kubernetes) and runtime environments (Docker).
  • Hands-on experience with Jupyter environments and GPU-accelerated AI frameworks (PyTorch, TensorFlow, Hugging Face).
  • Knowledge of Kubernetes and GPU computing for configuring and troubleshooting AI apps in shared environments.
  • Experience supporting researchers with model training, fine-tuning, inference, and optimization.
  • Proficiency in Python and shell scripting for automation and platform integration.
  • Familiarity with ML workflows and research-to-production transitions.
  • Strong problem-solving, communication, collaboration, and prioritization skills.
  • Ability to obtain Interim Secret, eventually Secret clearance; U.S. citizenship required.

Responsibilities

  • Own the AI platform configuration and evolution, adopting new capabilities and setting standards.
  • Collaborate with Linux/Kubernetes admins on deployments, upgrades, and integration.
  • Provide hands-on support to researchers using GPU-accelerated AI platforms for model work.
  • Troubleshoot AI workloads across Python, Jupyter, containers, Kubernetes, and GPU components.
  • Help scale workloads from single-GPU experiments to multi-GPU multi-node runs.
  • Develop reusable environments, workflows, automation, and documentation to boost reliability.

Skills

Python
Shell scripting
GPU scheduling
AI frameworks
Distributed training
Research collaboration
Problem solving

Education

Bachelor of Science in CS or related field

Tools

Docker
Kubernetes
Jupyter

Job description

Johns Hopkins Applied Physics Laboratory seeks an AI Solutions Engineer to bridge researchers and a scalable AI computing platform. You will configure and evolve the AI stack, collaborate on platform deployments, and assist researchers with GPU-accelerated workloads for training, inference, and deployment.

Ideal candidates have 3+ years in Kubernetes/Docker, hands-on Jupyter, and experience with PyTorch, TensorFlow, or similar frameworks.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

AI Platform Engineer: Scale GPU AI for Researchers
AI Platform Engineer: Scale GPU AI for Researchers

The Johns Hopkins University Applied Physics Laboratory • Laurel (MD)

On-site
USD 85,000 - 165,000
AI Platform Engineer: Scale GPU AI for Researchers
AI Platform Engineer: Scale GPU AI for Researchers

The Johns Hopkins University Applied Physics Laboratory • Laurel (MD)

On-site
USD 85,000 - 165,000
AI/ML Platform Engineer – Build Scalable AI Systems
AI/ML Platform Engineer – Build Scalable AI Systems

The Johns Hopkins University Applied Physics Laboratory • Laurel (MD)

On-site
USD 85,000 - 195,000
AI Infrastructure Architect: Kubernetes & GPU Scaling
AI Infrastructure Architect: Kubernetes & GPU Scaling

NVIDIA • United States

Remote
USD 272,000 - 431,000
Platform Software Engineer – AI Compute & Kubernetes
Platform Software Engineer – AI Compute & Kubernetes

ClearCompany Talent Management Software • Seattle (WA)

Hybrid
USD 147,000 - 183,000
AI/ML Platform Engineer: Kubernetes & GPU Infra at Scale
AI/ML Platform Engineer: Kubernetes & GPU Infra at Scale

Deepgram • San Francisco (CA)

On-site
USD 180,000 - 260,000
AI/ML-Driven HPC Simulation Engineer
AI/ML-Driven HPC Simulation Engineer

Rescale • United States

On-site
USD 120,000 - 180,000
Senior HPC & AI Software Engineer
Senior HPC & AI Software Engineer

The Johns Hopkins University • Baltimore (MD)

On-site
USD 80,000 - 120,000
AI Kernel / Cluster Engineer
AI Kernel / Cluster Engineer

Blue Signal Search • Santa Clara (CA)

On-site
USD 150,000 - 210,000
Senior GPU Cloud Infrastructure Engineer
Senior GPU Cloud Infrastructure Engineer

Jack & Jill • San Francisco (CA)

On-site
USD 140,000 - 190,000