An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Together AI, the AI Native Cloud, is hiring a software engineer to build a Kubernetes-native control plane for provisioning and running a GPU inference fleet. You will design a manifest-driven API where inference teams request clusters or models, with controllers handling reconciliation, provider selection, and lifecycle management.
The role emphasizes decoupling users from runtime complexity, improving utilization via defragmentation, bin-packing, and right-sizing, and owning the end-to-end
Together AI, the AI Native Cloud, is hiring a software engineer to build a Kubernetes-native control plane for provisioning and running a GPU inference fleet. You will design a manifest-driven API where inference teams request clusters or models, with controllers handling reconciliation, provider selection, and lifecycle management.
The role emphasizes decoupling users from runtime complexity, improving utilization via defragmentation, bin-packing, and right-sizing, and owning the end-to-end