Stand out for this role — generate a tailored resume and cover letter in about a minute.
RAPSYS TECHNOLOGIES PTE LTD is seeking an experienced HPC Systems Engineer to design, deploy, and maintain Linux-based HPC systems that support advanced computing workloads and research initiatives. You will work with researchers and enterprise workloads to ensure reliable, secure, and high-performance operations.
Key responsibilities include administering HPC clusters, storage and data management, capacity planning, and collaboration with software teams on AI/DL applications.
We are seeking an experienced System Engineer specializing in High Performance Computing (HPC) to join our dynamic team. You will play a key role in designing, deploying, and maintaining HPC systems to support advanced computational workloads and research initiatives.
Client is seeking an experienced HPC Systems Engineer (or Senior HPC Systems Engineer, depending on experience) to support and operate large-scale Linux-based high-performance computing (HPC), storage, and networking environments. This role supports research scientists, academic users, and enterprise workloads, ensuring reliable, secure, and high-performance HPC operations.
Administer, operate, and maintain Linux-based HPC clusters, including compute, storage, and high-speed networking.
Perform system monitoring, patching, upgrades, and capacity planning.
Troubleshoot and resolve hardware, software, OS, and network issues across HPC environments.
Participate in on-call or escalation support rotations as required.
Work with software engineers to support AI/DL applications and with desktop engineers to assist users as needed.
Provide advice and guidance to researchers on HPC application development, debugging, optimization, and parallelization.
Deliver HPC user training sessions and contribute to documentation and best-practice guides.
Meet all SLA requirements for incident and service request handling.
Comply with all policy and contract requirements.