Turn this role into an interview — a resume and cover letter built around what this employer wants.
Sobolev Research Center is seeking an ML Engineer to develop and optimize algorithms that accelerate large language model inference, directly impacting latency, cost efficiency, and scalability of production-grade AI systems.
You will explore and implement cutting-edge techniques such as speculative decoding, prompt compression, quantization, and generation optimisation to improve performance across deployment scenarios.
Sobolev Research Center is seeking an ML Engineer to develop and optimize algorithms that accelerate large language model inference, directly impacting latency, cost efficiency, and scalability of production-grade AI systems.
You will explore and implement cutting-edge techniques such as speculative decoding, prompt compression, quantization, and generation optimisation to improve performance across deployment scenarios.