AI Infrastructure Engineer

Mission.dev

Montreal (administrative region)

Hybrid

CAD 140,000 - 190,000

Full time

4 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Founding equity

Job summary

Mission.dev is seeking a founding Senior AI Infrastructure Engineer to design and operate large-scale AI workloads across public cloud and on-premises environments. You will be an individual contributor shaping secure, reliable platforms with deep systems expertise.

Reporting to the CTO, you will enable efficient model serving at scale, drive multi-cloud deployments, and collaborate with clients to integrate optimization solutions while visiting Montreal 2–3 times per quarter.

Qualifications

  • 5+ years in AI infrastructure or systems engineering.
  • Strong Python programming and API development experience.
  • Extensive Kubernetes, Helm, Ansible, Terraform expertise.
  • Proven track record shipping LLM serving systems.
  • Experience with Ray, RayClusters, or KubeRay.
  • Experience with agentic workflows.
  • Bachelor's degree in CS or related field.

Responsibilities

  • Design and manage multi-cloud and on-prem infrastructure for internal and external workloads.
  • Orchestrate large-scale deployments using infrastructure-as-code and container orchestration to support high user concurrency.
  • Benchmark AI inference workloads to identify performance bottlenecks and optimize hardware utilization.
  • Develop and maintain comprehensive monitoring systems to ensure high service quality and reliability.
  • Collaborate with customers to deploy and integrate optimization solutions within their existing infrastructure.
  • Build and maintain an internal laboratory for infrastructure testing and experimentation.

Skills

Python
API development
Distributed systems
Agentic workflows

Education

Bachelor's degree in Computer Science or related field

Tools

Kubernetes
Helm
Ansible
Terraform
Ray
RayClusters
KubeRay

Job description

Employment type: Full-time position, 40 hours per week, with direct employment by the client.

Location: Remote-first role for candidates based in Canada.

Travel requirement: Candidates must be available to visit the client's office in Montreal approximately two to three times per quarter.

Our company description

Mission.dev is the next-gen staffing platform for software talent.

We help you find, evaluate, and manage top software talent (contractors or direct hires) faster, smarter, and more efficiently.

Powered by AI. Backed by real humans.

About the client

An AI infrastructure company specializing in inference optimization for large-scale models.

The platform enhances token throughput and reduces latency by optimizing hardware utilization and decoding processes within a customer's cloud environment.

The technology focuses on improving the efficiency of model serving without requiring changes to existing infrastructure or code, ensuring data privacy while reducing operational costs for organizations deploying open-source models.

About the Role

As a founding Senior AI Infrastructure Engineer, you will report to the Chief Technology Officer to design and operate large-scale infrastructure for AI workloads on public cloud and on-premises environments. You will be an individual contributor with significant influence, combining software engineering with deep systems expertise to build secure and reliable platforms. Your work will focus on enabling efficient model serving at scale, ensuring the infrastructure can support a massive number of concurrent users while maintaining high service quality and performance.

What You'll Do
  • Design and manage multi-cloud and on-premises infrastructure for internal and external workloads.
  • Orchestrate large-scale deployments using infrastructure-as-code and container orchestration to support high user concurrency.
  • Benchmark AI inference workloads to identify performance bottlenecks and optimize hardware utilization.
  • Develop and maintain comprehensive monitoring systems to ensure high service quality and reliability.
  • Collaborate with customers to deploy and integrate optimization solutions within their existing infrastructure.
  • Build and maintain an internal laboratory for infrastructure testing and experimentation.
What You Bring
  • 5+ years of experience in AI infrastructure or systems engineering.
  • Proficiency in Python and experience building and operating scalable APIs.
  • Extensive experience with Kubernetes, Helm, Ansible, and Terraform.
  • Proven track record of shipping large language model (LLM) serving systems in production environments.
  • Experience with distributed computing frameworks such as Ray, RayClusters, or KubeRay.
  • Demonstrated experience with agentic workflows.
  • Bachelor's degree in Computer Science or a related technical field.
  • Founding-engineer equity and direct ownership.
  • Remote-friendly work environment. Visit the office 2-3x a quarter in Montreal.
  • Access to extensive compute resources and specialized hardware for experimentation.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AI/ML Infrastructure Engineer
AI/ML Infrastructure Engineer

BULL-IT SOLUTIONS LTD • Montreal

On-site
CAD 100,000 - 130,000
Senior LLMOps Engineer -Cloud / AI Infrastructure
Senior LLMOps Engineer -Cloud / AI Infrastructure

Talent To Hire Inc. • Toronto

On-site
CAD 120,000 - 160,000
Competitive salary
Meaningful equity
Innovative work culture
Forward Deployed Engineer
Forward Deployed Engineer

Robots & Pencils LP • Canada

On-site
CAD 176,000 - 244,000
Software Engineer, Data Infrastructure
Software Engineer, Data Infrastructure

Cohere • Montreal (administrative region)

Hybrid
CAD 90,000 - 120,000
Open and inclusive culture
Weekly lunch stipend and snacks
Full health and dental benefits
+4
Generative AI Engineer
Generative AI Engineer

Soho Square Solutions • Montreal (administrative region)

Hybrid
CAD 100,000 - 130,000
Applied AI Engineer
Applied AI Engineer

Worky • Montreal (administrative region)

Hybrid
CAD 70,000 - 110,000
Developer, AI Transformation Team
Developer, AI Transformation Team

Escalent • Toronto

On-site
CAD 135,000 - 155,000
System Software Engineer - AI
System Software Engineer - AI

Delos Data • Montreal (administrative region)

Hybrid
CAD 120,000 - 180,000
Meaningful equity
Benefits
401k
Software Engineer, Data Infrastructure
Software Engineer, Data Infrastructure

Talanto • Toronto, Montreal (administrative region)

Hybrid
CAD 120,000 - 180,000
Weekly lunch stipend
Health and dental benefits
Parental Leave top-up
+6
Full Stack Engineer
Full Stack Engineer

Randstad Enterprise • Montreal (administrative region)

On-site
CAD 90,000 - 140,000