Experience with HPC schedulers (e.g., YellowDog, Ray, Slurm, IBM Symphony)
Deep knowledge of Linux
Understanding of loosely and tightly coupled HPC workloads
Experience with large-scale systems
Monitoring and visualization of large-scale systems
Performance tuning of compute, network and storage
Knowledge of AWS cloud services
Automation and coding in Python
Strong analysis and problem-solving skills
Collaborative and pragmatic communication skills