AI/ML Infrastructure Engineer

BULL-IT SOLUTIONS LTD

Montreal

On-site

CAD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading IT solutions provider is seeking a Delivery Head in Montreal. The role focuses on managing large-scale systems while ensuring reliability and efficiency. Ideal candidates will possess strong programming fundamentals, experience in infrastructure operations, and excellent communication skills. This full-time position targets mid-senior level professionals and requires proficiency in modern technologies such as Docker and Kubernetes.

Skills

Production experience in SRE / Infrastructure / ops for large-scale systems
Strong programming/scripting skills (Python, Go, Java, or equivalent)
Deep experience with containerization (Docker), orchestration (Kubernetes)
Familiarity with GPU / AI compute clusters
Experience with monitoring / observability tools
Networking & systems engineering knowledge
Experience in capacity planning and incident response
Excellent communication and collaboration skills
Proven track record of reducing operational toil via automation

Job description

Delivery Head | Canada Recruitment | Talent Acquisition

Location: Montreal, Quebec, Canada

Seniority level: Mid-Senior level

Employment type: Full-time

Job function: Information Technology

Skills Required
  • Production experience in SRE / Infrastructure / ops for large-scale systems
  • Strong programming/scripting skills (Python, Go, Java, or equivalent)
  • Deep experience with containerization (Docker), orchestration (Kubernetes, etc.)
  • Familiarity with GPU / AI compute clusters, high-performance data storage, and distributed architectures
  • Experience with monitoring / observability / logging / alerting tools (Prometheus, Grafana, ELK / EFK, Datadog, etc.)
  • Networking & systems engineering knowledge (TCP/IP, DNS, routing, load balancing, distributed storage)
  • Solid experience in capacity planning, performance tuning, scaling, and incident response
  • Demonstrated ability to lead RCAs, deploy fixes, and drive reliability improvements
  • Experience in regulated environments (financial services, compliance, audit, security) is a strong plus
  • Excellent communication, documentation, and cross-team collaboration skills
  • Proven track record of reducing operational toil via automation
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI/ML Infrastructure Engineer — SRE for Scalable AI Clusters
AI/ML Infrastructure Engineer — SRE for Scalable AI Clusters

BULL-IT SOLUTIONS LTD • Montreal

On-site
CAD 100,000 - 130,000
Site Reliability Engineer, AI/ML Infrastructure
Site Reliability Engineer, AI/ML Infrastructure

Boson AI • Toronto

On-site
CAD 100,000 - 130,000
Senior Engineer-Cloud AI Infrastructure
Senior Engineer-Cloud AI Infrastructure

Huawei Canada • Markham

On-site
CAD 172,000 - 306,000
Data and AI Developer
Data and AI Developer

TEEMA • Montreal (administrative region)

On-site
CAD 90,000 - 130,000
AI Architect – T & I
AI Architect – T & I

Jobtailor • Montreal (administrative region)

On-site
CAD 150,000 - 210,000
Head Of Engineering/ Technical Leadership
Head Of Engineering/ Technical Leadership

Motion Recruitment Partners LLC • Toronto

On-site
CAD 150,000 - 210,000
Medical, Dental, and Vision Insurance
Vacation Time
Senior L3 AI Engineer (JS-3)
Senior L3 AI Engineer (JS-3)

Acestack • Montreal

Hybrid
CAD 120,000 - 170,000
Montréal [Hybrid] - DevOps - MLOps engineer
Montréal [Hybrid] - DevOps - MLOps engineer

QUANTEAM - North America (RAINBOW PARTNERS Group) • Montreal (administrative region)

On-site
CAD 95,000 - 150,000
AI Engineer
AI Engineer

Myticas Consulting • Toronto

On-site
CAD 90,000 - 130,000
Senior Engagement Manager - AI SaaS
Senior Engagement Manager - AI SaaS

Tech Talent International • Toronto

On-site
CAD 130,000 - 140,000