Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Trulyyy is seeking an experienced AI Cloud Network Operations Engineer in Singapore to operate and optimize high-performance AI/HPC network infrastructure across data centers. The role emphasizes incident handling, performance troubleshooting, and scalable network configuration for large GPU clusters.
Responsibilities include AI network operations, fabric optimization with InfiniBand/RoCEv2, and automation using Python/Go; familiarity with Prometheus/Grafana/Zabbix is expected.
Our client is a global technology company operating large-scale AI/HPC and data center infrastructure across multiple international markets. As its AI cloud capabilities continue to scale, the company is looking for an experienced AI Cloud Network Operations Engineer to operate and optimize high-performance network infrastructure supporting large-scale GPU computing environments.