An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Luma AI seeks a systems engineer to own large-scale inference deployments. You will integrate new architectures into the inference engine, scale fleets across thousands of machines, and keep GPU pools busy while meeting SLOs.
This role emphasizes model serving, scheduling, and reliability at scale. Responsibilities include building tooling to measure and optimize inference workloads, collaborating across research and infra teams, and maintaining CI/CD for model checkpoints and SDKs.
Luma AI seeks a systems engineer to own large-scale inference deployments. You will integrate new architectures into the inference engine, scale fleets across thousands of machines, and keep GPU pools busy while meeting SLOs.
This role emphasizes model serving, scheduling, and reliability at scale. Responsibilities include building tooling to measure and optimize inference workloads, collaborating across research and infra teams, and maintaining CI/CD for model checkpoints and SDKs.