HPC Customer Solutions Engineer

GTN Technical Staffing

United States

On-site

USD 120,000 - 180,000

Full time

29 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

GTN Technical Staffing is seeking an HPC Customer Solutions Engineer to lead the design, integration, and delivery of HPC and AI infrastructure solutions. The role focuses on scalable architectures across GPU/CPU compute, storage, networking, Kubernetes, orchestration, and security.

The ideal candidate will have deep HPC and AI infrastructure expertise, strong hands-on system design and performance tuning, and the ability to translate customer requirements into production-ready solutions.

Qualifications

  • Deep HPC and AI infrastructure expertise.
  • Strong system design and performance tuning experience.
  • Experience translating customer requirements into scalable, production-ready solutions.

Responsibilities

  • Understand workload requirements, performance targets, and objectives.
  • Lead technical discovery sessions focusing on bottlenecks, scalability, and infra needs.
  • Serve as trusted technical advisor throughout the solution lifecycle.
  • Design end-to-end HPC and AI architectures across compute, storage, networking, orchestration, and security.
  • Recommend hardware and software solutions aligned with performance, scalability, reliability, and efficiency goals.
  • Develop architecture blueprints, integration plans, and technical documentation.
  • Design solutions supporting GPU-intensive AI/ML, LLM, and advanced compute workloads.

Skills

GPU architectures
NVIDIA CUDA
Slurm
Kubernetes
InfiniBand/RDMA
Storage: Lustre/GPFS
Security & encryption
Linux performance
Customer-facing comms
Architecture design

Education

Bachelor's or Master's in CS/Engineering/Physics

Job description

Compensation: Competitive Base Salary + Performance Bonus

Overview

Our client is seeking an HPC Customer Solutions Engineer to lead the technical design, integration, and delivery of high-performance computing and AI infrastructure solutions.

This is a highly technical, customer-facing role focused on designing scalable architectures across GPU/CPU compute, storage, networking, Kubernetes, orchestration, and security. The position spans the full solution lifecycle from technical discovery and workload analysis through proof-of-concept, deployment, and ongoing optimization.

The ideal candidate brings deep HPC and AI infrastructure expertise, strong hands-on system design and performance tuning experience, and the ability to translate complex customer requirements into scalable, production-ready solutions.

Key Responsibilities
Customer Engagement & Technical Discovery
  • Work directly with customers to understand workload requirements, performance targets, and technical objectives.
  • Lead technical discovery sessions focused on application behavior, bottlenecks, scalability, and infrastructure requirements.
  • Serve as a trusted technical advisor throughout the solution lifecycle.
  • Design end-to-end HPC and AI architectures across compute, storage, networking, orchestration, and security.
  • Recommend hardware and software solutions aligned with performance, scalability, reliability, and efficiency goals.
  • Develop architecture blueprints, integration plans, and technical documentation.
  • Design solutions supporting GPU-intensive AI/ML, LLM, and advanced compute workloads.
Performance & Workload Optimization
  • Support proof-of-concept, benchmarking, and performance-validation initiatives.
  • Perform workload profiling, system tuning, and infrastructure optimization.
  • Identify bottlenecks across compute, storage, networking, and orchestration layers.
  • Recommend improvements that increase workload performance, scalability, and resilience.
Implementation & Delivery
  • Provide technical leadership during deployment and integration.
  • Partner with customers and internal Engineering, Product, and Operations teams throughout implementation.
  • Support solutions from architecture through production deployment and optimization.
  • Troubleshoot complex infrastructure and workload issues during delivery.
Technical Leadership
  • Maintain expertise across emerging HPC, AI, GPU, storage, networking, and orchestration technologies.
  • Build relationships with technology partners across GPU, networking, and storage ecosystems.
  • Contribute to reference architectures, reusable design patterns, and technical best practices.
  • Lead customer workshops, architecture reviews, and technical presentations.
Required Qualifications
  • Strong technical expertise across:
  • GPU and CPU architectures
  • NVIDIA / CUDA ecosystem
  • Slurm and Kubernetes
  • InfiniBand, RDMA, and RoCE
  • Lustre, GPFS / Spectrum Scale, Ceph, VAST, or similar storage platforms
  • Kubernetes and container orchestration
  • Identity, encryption, and infrastructure security
  • Strong Linux systems knowledge, including tuning and performance analysis.
  • Experience translating workload requirements into detailed technical architectures.
  • Experience with proof-of-concept, benchmarking, or workload optimization.
  • Strong customer-facing communication and presentation skills.
  • Ability to work effectively with engineering, product, operations, and executive stakeholders.
Preferred Experience
  • AI/ML, LLM, GPU, or HPC workloads.
  • NVIDIA GPU infrastructure.
  • Automation and Infrastructure-as-Code.
  • Workload migration and performance engineering.
  • Next-generation GPU and high-speed interconnect technologies.
  • Bachelor's or Master's degree in Computer Science, Engineering, Physics, or related field.
  • Relevant cloud, Linux, networking, Kubernetes, or security certifications.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Solutions Engineer, AI Infrastructure
Senior Solutions Engineer, AI Infrastructure

VAST Data • New York (NY)

On-site
USD 150,000 - 200,000
HPC & AI Infra Solutions Engineer
HPC & AI Infra Solutions Engineer

GTN Technical Staffing • United States

On-site
USD 120,000 - 180,000
Manager of HPC Solution Architect
Manager of HPC Solution Architect

Coda Search│Staffing • Dallas (TX)

On-site
USD 180,000 - 230,000
Relocation package available
HPC Customer Solutions Engineer
HPC Customer Solutions Engineer

NextSilicon • Minneapolis (MN)

On-site
USD 120,000 - 180,000
Solutions Architect
Solutions Architect

Addison Group • Dallas (TX)

On-site
USD 180,000 - 260,000
Full medical, dental, vision
401(k)
PTO 25 days
+2
Sr HPC Hardware Engineer
Sr HPC Hardware Engineer

Career Techniques • Dallas (TX)

On-site
USD 120,000 - 180,000
HPC Customer Solutions Engineer
HPC Customer Solutions Engineer

NextSilicon • Austin (TX)

On-site
USD 120,000 - 180,000
HPC Platform Engineer
HPC Platform Engineer

Addison Group • Dallas (TX)

Hybrid
USD 180,000 - 260,000
Medical, dental, and vision insurance
401(k)
25 days PTO
+3
HPC Performance and Validation Engineer
HPC Performance and Validation Engineer

Addison Group • Dallas (TX)

On-site
USD 180,000 - 260,000
100% paid medical, dental, vision
401(k)
25 days PTO
+3
HPC Solution Architect
HPC Solution Architect

Coda Search│Staffing • Dallas (TX)

On-site
USD 120,000 - 160,000