A technology company seeks an experienced candidate to manage and optimize job scheduling systems in a large-scale environment. This role involves analyzing performance data, resolving service-impacting issues, and implementing automation. Ideal candidates will have a Bachelor's degree in Computer Science and over 5 years of experience with Linux-based infrastructures and scheduling systems like LSF and Slurm. NVIDIA offers competitive salaries, equity, and comprehensive benefits.
Qualifications
5+ years of experience operating large-scale Linux-based compute infrastructure.
Experience with job scheduling systems (LSF, Slurm) in HPC or silicon design.
Responsibilities
Manage job scheduling systems in a multi-site environment.
Analyze performance data to improve utilization and throughput.
Implement automation to reduce manual effort.
Skills
Linux systems administration
Job scheduling systems (LSF, Slurm)
Problem-solving skills
Communication skills
Education
Bachelor’s degree in Computer Science or related field
Tools
CentOS/RHEL
Docker
Slurm
LSF
Job description
A technology company seeks an experienced candidate to manage and optimize job scheduling systems in a large-scale environment. This role involves analyzing performance data, resolving service-impacting issues, and implementing automation. Ideal candidates will have a Bachelor's degree in Computer Science and over 5 years of experience with Linux-based infrastructures and scheduling systems like LSF and Slurm. NVIDIA offers competitive salaries, equity, and comprehensive benefits.