A leading AI solutions company is hiring an ML Engineer in Toronto, Canada. In this role, you will design and maintain serving stacks for machine learning models and automate training pipelines using Kubernetes. Candidates should have solid experience with Python and knowledge of ML infrastructures. Join a team committed to deploying innovative AI technologies and enjoy perks like medical benefits, paid parental leave, and monthly wellness allowances.
Qualifications
5+ years writing production software; 2+ years focused on ML platform or infra.
Proven experience with one or more serving frameworks (e.g., vLLM, Triton, TorchServe).
Culture of rigorous testing, code review, and continuous delivery.
Responsibilities
Design, build, and maintain low-latency serving stacks for ML models.
Orchestrate data prep, training, and evaluation workflows on Kubernetes.
Profile and tune throughput, memory, and cost efficiently.
Skills
Python (async, typing, packaging, performance)
Golang (systems components)
Distributed systems
Kubernetes
Tools
Kubernetes
Terraform
Triton
Job description
A leading AI solutions company is hiring an ML Engineer in Toronto, Canada. In this role, you will design and maintain serving stacks for machine learning models and automate training pipelines using Kubernetes. Candidates should have solid experience with Python and knowledge of ML infrastructures. Join a team committed to deploying innovative AI technologies and enjoy perks like medical benefits, paid parental leave, and monthly wellness allowances.