Get more replies from employers
Send a job-specific resume in minutes.
Luma AI in San Francisco is building cutting-edge multimodal AI systems. We seek a seasoned Platform Engineer to ship new model architectures into our inference engine and optimize deployments across clusters.
You will collaborate with research, engineering and infra, develop scalable scheduling, CI/CD pipelines, and tooling to measure and ensure uptime for inference workloads across thousands of GPUs. Proficiency in Python, Linux, Docker, Kubernetes, and model deployment frameworks is required;
Luma's mission is to build multimodal AI to expand human imagination and capabilities. We believe that multimodality is critical for intelligence. To go beyond language models and build more aware, capable and useful systems, the next step function change will come from vision. So we are working on training and scaling up multimodal foundation models for systems that can see and understand, show and explain, and eventually interact with our world to effect change.