Stand out for this role — generate a tailored resume and cover letter in about a minute.
Get past ATS filters
Benefits offered by this job
Medical, dental, and vision plans
Paid parental leave
Monthly health & wellness allowance
Work-from-home office stipend
Lunch reimbursement
PTO: 3 weeks in Canada
Job summary
A leading AI solutions company is hiring an ML Engineer in Toronto, Canada. In this role, you will design and maintain serving stacks for machine learning models and automate training pipelines using Kubernetes. Candidates should have solid experience with Python and knowledge of ML infrastructures. Join a team committed to deploying innovative AI technologies and enjoy perks like medical benefits, paid parental leave, and monthly wellness allowances.
Qualifications
5+ years writing production software; 2+ years focused on ML platform or infra.
Proven experience with one or more serving frameworks (e.g., vLLM, Triton, TorchServe).
Culture of rigorous testing, code review, and continuous delivery.
Responsibilities
Design, build, and maintain low-latency serving stacks for ML models.
Orchestrate data prep, training, and evaluation workflows on Kubernetes.
Profile and tune throughput, memory, and cost efficiently.
Skills
Python (async, typing, packaging, performance)
Golang (systems components)
Distributed systems
Kubernetes
Tools
Kubernetes
Terraform
Triton
Job description
A leading AI solutions company is hiring an ML Engineer in Toronto, Canada. In this role, you will design and maintain serving stacks for machine learning models and automate training pipelines using Kubernetes. Candidates should have solid experience with Python and knowledge of ML infrastructures. Join a team committed to deploying innovative AI technologies and enjoy perks like medical benefits, paid parental leave, and monthly wellness allowances.