GPU Infrastructure Engineer for AI Inference & Serving

P2P

Montreal (administrative region)

On-site

CAD 90,000 - 120,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

A diversified trading firm is seeking an HPC Specialist to join their AI and Multi Asset Systematic Strategies team. This role focuses on building and managing GPU infrastructure for AI and ML workloads. Candidates should have strong experience in DevOps, GPU technology, and Kubernetes orchestration. A Bachelor’s or Master’s degree in a related field is required, along with a proven track record in optimizing deep learning workloads. This position is based in Montreal, Quebec.

Qualifications

  • 5+ years in DevOps, SRE, or infrastructure engineering roles.
  • Strong experience with model serving frameworks and GPU driver management.
  • Hands-on experience optimizing deep learning workloads on GPU clusters.

Responsibilities

  • Deploy, maintain, and optimize GPU infrastructure for large-scale LLM inference workloads.
  • Manage GPU-enabled Kubernetes clusters for LLM and ML workloads.
  • Collaborate with ML engineers on model performance and inference acceleration.

Skills

GPU infrastructure
Kubernetes orchestration
Python
Bash scripting
Deep learning workloads
Distributed systems
Monitoring tools

Education

Bachelor's or Master's degree in Computer Science

Tools

Ansible
Terraform
Prometheus
Grafana
vLLM
SGLang

Job description

A diversified trading firm is seeking an HPC Specialist to join their AI and Multi Asset Systematic Strategies team. This role focuses on building and managing GPU infrastructure for AI and ML workloads. Candidates should have strong experience in DevOps, GPU technology, and Kubernetes orchestration. A Bachelor’s or Master’s degree in a related field is required, along with a proven track record in optimizing deep learning workloads. This position is based in Montreal, Quebec.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

GPU Infra Architect for AI Inference
GPU Infra Architect for AI Inference

DRW Holdings, LLC. • Montreal (administrative region)

On-site
CAD 100,000 - 130,000
Remote: Technical Lead, GPU Infra & HPC Platform
Remote: Technical Lead, GPU Infra & HPC Platform

Jobgether • Canada

Remote
CAD 150,000 - 230,000
100% remote position
Lead architecture and delivery of GPU/
International and distributed team
Remote GPU Cloud Platform Engineer for Multi-Cluster AI
Remote GPU Cloud Platform Engineer for Multi-Cluster AI

Yotta Labs • Canada

Remote
CAD 90,000 - 120,000
AI Infrastructure Engineer — GPU Clusters & HPC
AI Infrastructure Engineer — GPU Clusters & HPC

Veeda AI • Toronto

On-site
CAD 120,000 - 170,000
HPC Specialist
HPC Specialist

DRW Holdings, LLC. • Montreal (administrative region)

On-site
CAD 100,000 - 130,000
Senior SRE: AI/ML HPC Infra & GPU Cluster
Senior SRE: AI/ML HPC Infra & GPU Cluster

Boson AI • Toronto

On-site
CAD 100,000 - 130,000
Senior AI Inference Systems Engineer
Senior AI Inference Systems Engineer

NVIDIA Corporation • Toronto

Hybrid
CAD 170,000 - 275,000
Equity
Benefits
HPC Specialist
HPC Specialist

P2P • Montreal (administrative region)

On-site
CAD 90,000 - 120,000
Site Reliability Engineer, AI/ML Infrastructure
Site Reliability Engineer, AI/ML Infrastructure

Boson AI • Toronto

On-site
CAD 100,000 - 130,000
Senior Software Engineer, AI Inference Systems
Senior Software Engineer, AI Inference Systems

NVIDIA Corporation • Toronto

On-site
CAD 170,000 - 275,000
Equity
Benefits