An application made for this job — a tailored resume and cover letter that speak straight to the posting.
ByteDance is seeking a Research Engineer/Scientist to design and implement efficient models for large-scale generative AI, with emphasis on distillation, compression, and deployment. You will work on model acceleration and hardware-aware inference within the Vision-Applied Research team in San Jose.
The role requires experience in quantization, diffusion/ autoregressive methods, and a strong background in PyTorch or JAX to optimize model performance and scalability for real-world applications.
ByteDance is seeking a Research Engineer/Scientist to design and implement efficient models for large-scale generative AI, with emphasis on distillation, compression, and deployment. You will work on model acceleration and hardware-aware inference within the Vision-Applied Research team in San Jose.
The role requires experience in quantization, diffusion/ autoregressive methods, and a strong background in PyTorch or JAX to optimize model performance and scalability for real-world applications.