Get more replies from employers
Send a job-specific resume in minutes.
MRE Consulting in Houston seeks a HPC AI Systems Administrator to architect a secure, scalable on-prem HPC compute platform, enabling your development teams to fine-tune and deploy production ML models. You will lead deployment, manage multi-GPU hardware, and oversee Linux, GPU drivers, CUDA, NCCL, containers, and orchestration tools, ensuring performance and security.
This role requires 3+ years in HPC or enterprise GPU infra, strong Linux, and experience with Slurm/Kubernetes, InfiniBand
MRE Consulting in Houston seeks a HPC AI Systems Administrator to architect a secure, scalable on-prem HPC compute platform, enabling your development teams to fine-tune and deploy production ML models. You will lead deployment, manage multi-GPU hardware, and oversee Linux, GPU drivers, CUDA, NCCL, containers, and orchestration tools, ensuring performance and security.
This role requires 3+ years in HPC or enterprise GPU infra, strong Linux, and experience with Slurm/Kubernetes, InfiniBand