An application made for this job — a tailored resume and cover letter that speak straight to the posting.
Together AI in San Francisco is seeking a research intern to accelerate inference systems for large foundation models. You will explore distributed inference, compiler-aware optimization, and cross-layer strategies across models, systems, and hardware.
Join a collaborative team, publish findings, and contribute to faster serving and scalable deployment, with a 12–14 week program from January through April, housing stipends, and competitive hourly compensation.
Together AI in San Francisco is seeking a research intern to accelerate inference systems for large foundation models. You will explore distributed inference, compiler-aware optimization, and cross-layer strategies across models, systems, and hardware.
Join a collaborative team, publish findings, and contribute to faster serving and scalable deployment, with a 12–14 week program from January through April, housing stipends, and competitive hourly compensation.