Senior SRE

CloudFactory Limited

Canada

Hybrid

CAD 120,000 - 160,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid Working Model
Comprehensive medical cover
Group life insurance
Quarterly variable compensation
Market competitive salary
Personal development and growth

Job summary

CloudFactory Limited in Canada seeks a Senior SRE to design and build scalable infrastructure and pipelines supporting automation, reliability, and performance of production systems. You will apply site-reliability practices with autonomy while clearly communicating complex issues to stakeholders.

This full-time, fixed-term 6-month role requires 5+ years in production infra, strong Python, Docker/Kubernetes, Terraform, and cloud experience (GCP/AWS). Hybrid work model offered.

Qualifications

  • 5+ years of experience building and operating infrastructure in production environments.
  • Fluent in Python, with strong production-ready coding experience.
  • Experience with Docker and Kubernetes.
  • Knowledgeable about cloud platforms such as GCP or AWS.
  • Experience using Infrastructure as Code (IaC) tools such as Terraform.
  • Experience using CI/CD platforms to automate build, test, and deployment pipelines.
  • Comfortable applying site-reliability principles across production systems.

Responsibilities

  • Design and implement new core infrastructure components with autonomy.
  • Optimize and improve deployment pipelines, environment provisioning, and batch jobs.
  • Use IaC tools like Terraform to manage and scale infrastructure.
  • Develop CI/CD pipelines to automate build, test, deployment, and monitoring processes.
  • Set up monitoring, alerting, and observability tooling for system health.
  • Collaborate with engineers, product, and stakeholders on infrastructure delivery.

Skills

SRE
Python
Docker
Kubernetes
Cloud (GCP/AWS)
Terraform
CI/CD
Observability
Automation

Education

CS degree / Engineering degree
Equivalent practical experience

Tools

Terraform
Docker
Kubernetes
Prometheus
Grafana
Ansible

Job description

At CloudFactory, we are a mission-driven team passionate about unlocking the potential of AI to transform the world. By combining advanced technology with a global network of talented people, we make unusable data usable, driving real-world impact at scale.

More than just a workplace, we’re a global community founded on strong relationships and the belief that meaningful work transforms lives. Our commitment to earning, learning, and serving fuels everything we do as we strive to connect one million people to meaningful work and build leaders worth following.

Our Culture

At CloudFactory, we believe in building a workplace where everyone feels empowered, valued, and inspired to bring their authentic selves to work. We are:

  • Mission-Driven: We focus on creating economic and social impact.
  • People-Centric: We care deeply about our team’s growth, well-being, and sense of belonging.
  • Innovative: We embrace change and find better ways to do things together.
  • Globally Connected: We foster collaboration between diverse cultures and perspectives.

If you’re passionate about innovation, collaboration, and making a real impact, we’d love to have you on board!

Role Summary

As a Senior SRE, you will design and build scalable infrastructure, working closely with cross-functional teams to develop systems and pipelines that support the automation, reliability, and scalability of our production environments. You will bring a high degree of autonomy to designing new infrastructure components and applying site-reliability practices across our systems, while communicating complex technical issues clearly to stakeholders across the business. This is an exciting opportunity to make a real impact while working alongside talented people from developing nations.

Please note: This is a full-time, fixed-term employee position with an expected duration of 6 months.

Responsibilities
Infrastructure design and automation
  • Design and implement new core infrastructure components with a high degree of autonomy.
  • Optimize and improve existing systems and operations, such as deployment pipelines, environment provisioning, and high-throughput batch jobs.
  • Use Infrastructure as Code (IaC) tools, such as Terraform, to manage and scale complex infrastructure.
CI/CD automation
  • Develop CI/CD pipelines to automate build, test, deployment, and monitoring processes.
  • Create and manage multi-step CI/CD pipelines, including environment setup and artifact handling.
Reliability and availability
  • Support the reliability, availability, and performance of production systems, applying site-reliability practices across the infrastructure.
  • Set up monitoring, alerting, and observability tooling to maintain visibility into system health.
Collaboration and communication
  • Collaborate closely with software engineers, product, and business stakeholders on the design and delivery of infrastructure and deployment systems.
  • Communicate complex technical issues clearly to stakeholders from technical and non-technical backgrounds alike.
Must-have skills (required)
  • 5+ years of experience building and operating infrastructure in production environments.
  • Fluent in Python, with strong experience writing production-ready code.
  • Experience with Docker and Kubernetes.
  • Knowledgeable about cloud platforms such as GCP or AWS.
  • Experience using Infrastructure as Code (IaC) tools such as Terraform.
  • Experience using CI/CD platforms to automate build, test, and deployment pipelines.
  • Comfortable applying site-reliability principles, such as availability, observability, and automation, across production systems.
Academic and professional requirements
  • Degree in Computer Science, Engineering, or another quantitative or computational field, or equivalent practical experience.
Nice-to-have skills (preferred)
  • Familiarity with monitoring tools such as Prometheus or Grafana.
  • Experience with configuration management tools (e.g., Ansible, Chef, Puppet).
  • Exposure to multi-cloud or hybrid-cloud environments.
  • Great Mission and Culture
  • Meaningful Work
  • Market competitive salary
  • Quarterly variable compensation
  • Hybrid Working Model
  • Comprehensive medical cover
  • Group life insurance
  • Personal development and growth opportunities

At CloudFactory, we believe that work should be more than just a job—it should be a platform for growth, impact, and community. Here, you’ll earn with purpose, learn every day, and serve a mission that truly matters. If you're looking for a career where you can develop professionally, contribute meaningfully, and be part of a global movement, we’d love to have you on this journey!

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Developer
Senior Site Reliability Developer

United States Digital Space LLC • Toronto

On-site
CAD 107,000 - 157,000
Salary transparency
In-person onboarding
Site Reliability Engineer
Site Reliability Engineer

Future Secure AI • Toronto

On-site
CAD 90,000 - 120,000
Flexible work environment
Competitive salary
Diversity and creativity
Senior Site Reliability Engineer
Senior Site Reliability Engineer

iManage • Toronto

Hybrid
CAD 90,000 - 120,000
Market-competitive salary
Annual performance-based bonus
Comprehensive Health, Vision, Dental, and Life insurance
+4
Senior Software Developer, DevOps & Infrastructure
Senior Software Developer, DevOps & Infrastructure

OSEDEA • Montreal (administrative region)

Hybrid
CAD 85,000 - 120,000
Competitive Salary and RRSP contribution
Flexible hours
Work from anywhere for up to 8 weeks
+2
Senior DevOps Engineer
Senior DevOps Engineer

Quest Global • Vancouver

On-site
CAD 100,000 - 120,000
401(k) matching
Health insurance
Dental insurance
+5
Senior DevOps
Senior DevOps

Quartermaster inc. • Toronto

Hybrid
CAD 160,000 - 215,000
30 days of PTO annually
Health, dental, and wellness benefits
Tech allowance benefit
+1
Staff Platform Engineer
Staff Platform Engineer

Robots and Pencils • Calgary

Hybrid
CAD 96,000 - 138,000
Senior Staff Software Developer, Developer Productivity & AI Tooling
Senior Staff Software Developer, Developer Productivity & AI Tooling

United States Digital Space LLC • Toronto

On-site
CAD 130,000 - 170,000
Performance driven compensation
Supplemental health insurance
Mental health support programs
+1
Senior Site Reliability Engineer - Global Infra & CI/CD Impact
Senior Site Reliability Engineer - Global Infra & CI/CD Impact

CloudFactory Limited • Canada

Hybrid
CAD 120,000 - 160,000
Hybrid Working Model
Comprehensive medical cover
Group life insurance
+3
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Rootly • Toronto

On-site
CAD 120,000 - 180,000
Competitive compensation
Comprehensive medical coverage
3 weeks of vacation
+3