Get more replies from employers
Send a job-specific resume in minutes.
Our client, a well-funded AI company, designs and runs large-scale compute infrastructure powering frontier model training and inference. They seek an infrastructure engineer to design, operate, and improve their GPU cluster platform, a deeply technical role at the junction of distributed systems and ML platform engineering.
You will own the compute platform, optimize GPU resource sharing, tackle bottlenecks across compute, storage and networking, and build automation and observability.
Our client, a well-funded AI company, designs and runs large-scale compute infrastructure powering frontier model training and inference. They seek an infrastructure engineer to design, operate, and improve their GPU cluster platform, a deeply technical role at the junction of distributed systems and ML platform engineering.
You will own the compute platform, optimize GPU resource sharing, tackle bottlenecks across compute, storage and networking, and build automation and observability.