An application made for this job — a tailored resume and cover letter that speak straight to the posting.
GMI Cloud is seeking an Infrastructure Backend Engineer to design, build, and maintain scalable AI infrastructure in Mountain View, CA. The role emphasizes cloud computing, distributed systems, and DevOps practices to enable efficient AI infrastructure operations.
You will contribute to large-scale training and inference, automate resource provisioning, and implement telemetry with Prometheus, Grafana, and Mimir while staying current with GPU technology.
GMI Cloud is a fast-growing, AI-native infrastructure company delivering high-performance GPU compute, inference services, and infrastructure for AI agents. Following 8x ARR growth, GMI Cloud continues to scale rapidly across the U.S. and APAC. As a Reference Platform NVIDIA Cloud Partner (NCP) and a validated leading NCP across both markets, we power production AI for leading AI-native companies including Fireworks AI, Cartesia, Reflection, and OpenRouter. From large-scale compute to optimized inference and agentic workloads, GMI Cloud gives AI teams the infrastructure they need to build, deploy, and scale on one unified cloud. One cloud for compute, inference, and agents.
We are seeking a talented and highly skilled Infrastructure Backend Engineering Development Engineer to design, build, and maintain the scalable infrastructure that supports GMI AI/ML initiatives. The ideal candidate will have a strong background in cloud computing, distributed systems, and DevOps practices to enable efficient AI infrastructure operations.
Meeting every qualification is not required—if you’re excited about this role, we’d love to hear from you. We believe diverse perspectives and experiences strengthen our team.