Site Reliability Engineer

Future Secure AI

Toronto

On-site

CAD 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Flexible work environment
Competitive salary
Diversity and creativity

Job summary

Future Secure AI in Toronto is looking for a Site Reliability Engineer to design and operate platforms powering AI Co-Workers. This role requires 5+ years of experience with Kubernetes and Terraform.

You'll build reliable infrastructure, implement automation, and partner closely with engineering teams. Enjoy a competitive salary, flexible work environment, and a chance to work with the best in the field.

Qualifications

  • 5+ years of relevant work experience as a Site Reliability Engineer or similar.
  • Experience working with Kubernetes on major cloud providers like AWS or Azure.
  • Proficiency in at least two programming or scripting languages.

Responsibilities

  • Design, build, and operate reliable production infrastructure for AI.
  • Own Kubernetes platforms for AI workloads.
  • Build and maintain infrastructure as code using Terraform.

Skills

Kubernetes
Terraform
Python
DevOps
AWS

Job description

At Future Secure AI, we're building something genuinely new — and we're looking for people bold enough to build it with us. We work at the frontier of AI, tackling big, real-world problems for global enterprises across multiple industries, armed with state‑of‑the‑art technology and a culture that prizes courage, rigor, and relentless curiosity. Our BRAVER values aren't just words on a wall — they describe the kind of people we are and the standard we hold ourselves to every day. Our leadership team is entrepreneurial, experienced, and accessible, with an open‑door policy that means you'll never be just a number here. We invest seriously in your growth because we know our success depends on yours. If you're ready to work alongside some of the brightest minds in the industry, push into uncharted territory, and do work that genuinely matters, Future Secure AI is the place for you.

About the Role

We are looking for a Site Reliability Engineer to help design, build, and operate the platforms that power AI Co‑Workers. This is a hands‑on role for an engineer who enjoys owning reliability end‑to‑end and working closely with product, AI, and engineering teams.

Responsibilities
  • Design, build, and operate reliable production infrastructure supporting AI Co‑Workers
  • Own Kubernetes‑based platforms used to deploy and run AI workloads
  • Build and maintain infrastructure as code using Terraform
  • Implement and maintain Helm‑based deployment workflows
  • Define, measure, and improve system reliability using SLIs, SLOs, and SLAs
  • Participate in on‑call rotation, incident response, root cause analysis, and post‑mortems
  • Reduce operational toil through automation and engineering improvements
  • Build and improve observability across monitoring, logging, and alerting
  • Partner closely with engineers to ensure systems are resilient, scalable, and secure
  • Operate across build, deploy, and operate phases of the software lifecycle
Minimum Qualifications
  • 5+ years of relevant work experience as a Site Reliability Engineer, DevOps Engineer or similar role
  • Hands on Kubernetes experience designing, building or operating workloads on EKS, AKS, GKE or self‑managed Kubernetes
  • Terraform experience for infrastructure provisioning and automation
  • Experience with Help for Kubernetes application deployment
  • Hands on experience working with at least one major cloud provide such as AWS, Azure or Google Cloud
  • Experience with at least two programming or scripting languages such as Python, Go, Java, Bash, Ruby or PowerShell
  • Experience in reliability engineering, on‑call rotations, incident response, post‑mortems and toil reduction
Preferred Qualifications
  • Experience working within a defined SDLC, including CI/CD, release processes, and end‑to‑end delivery from design to operations
  • Experience with ArgoCD or GitOps‑style deployment approaches
  • DevOps or DevSecOps experience, including CI/CD ownership, infrastructure automation, and security considerations
  • Relevant certifications such as CKA, CKAD, cloud certifications, DevOps, DevSecOps, or programming credentials
Why Join Us?
  • A high‑performance culture
  • State‑of‑the‑art technology
  • Experience world‑class leadership
  • Scale of impact and purpose
  • A competitive salary and a huge growth trajectory
  • Work with the best in the industry
  • Flexible work environment
  • Diversity and creativity
Disclaimer

We do not wish to be contacted by recruitment agencies. Our hiring process is managed in‑house and the best way for candidates to express interest is by applying with your resume through our company website.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Solutions Architect
Solutions Architect

Future Secure AI • Toronto

On-site
CAD 90,000 - 150,000
Flexible work environment
Growth opportunities
Engagement Lead Toronto, CAN
Engagement Lead Toronto, CAN

Future Secure AI Pty • Toronto

On-site
CAD 100,000 - 130,000
Competitive salary
Performance incentives
Equity participation
+2
Quality Assurance Engineer
Quality Assurance Engineer

Future Secure AI • Toronto

On-site
CAD 75,000 - 95,000
Competitive salary
Flexible work environment
State-of-the-art technology
+1
Software Engineer
Software Engineer

FutureFit • Canada

On-site
USD 150,000 - 185,000
Staff Platform Engineer
Staff Platform Engineer

Robots and Pencils • Calgary

Hybrid
CAD 96,000 - 138,000
Senior SRE
Senior SRE

CloudFactory Limited • Canada

Hybrid
CAD 120,000 - 160,000
Hybrid Working Model
Comprehensive medical cover
Group life insurance
+3
Site Reliability Engineer (SRE) – UI/UX
Site Reliability Engineer (SRE) – UI/UX

Software Mind • Montreal (administrative region)

Hybrid
CAD 110,000 - 165,000
Competitive salary
Laptop provided
Professional development
+2
Senior Site Reliability Engineer (SRE) – Kubernetes
Senior Site Reliability Engineer (SRE) – Kubernetes

Software Mind • Montreal (administrative region)

Hybrid
CAD 120,000 - 170,000
Competitive salary
Laptop provided
Professional development
+2
Senior Site Reliability Developer
Senior Site Reliability Developer

United States Digital Space LLC • Toronto

On-site
CAD 107,000 - 157,000
Salary transparency
In-person onboarding
Senior Site Reliability Engineer
Senior Site Reliability Engineer

AlleyCorp • Canada

On-site
CAD 197,000 - 225,000
Stock options
Health benefits
Unlimited PTO
+2