Turn this role into an interview — a resume and cover letter built around what this employer wants.
Innowise is seeking a talented ML/LLM production deployment engineer in Poland to deploy and optimize models for production inference across cloud GPU and edge targets. You will work with leading inference frameworks and implement optimization techniques to ensure latency, throughput, and memory efficiency.
The role requires strong Python, ML/LLM fundamentals, and hands-on experience with at least one inference framework; knowledge of transformer architectures is essential for scalable
Innowise is seeking a talented ML/LLM production deployment engineer in Poland to deploy and optimize models for production inference across cloud GPU and edge targets. You will work with leading inference frameworks and implement optimization techniques to ensure latency, throughput, and memory efficiency.
The role requires strong Python, ML/LLM fundamentals, and hands-on experience with at least one inference framework; knowledge of transformer architectures is essential for scalable