Site Reliability Engineer

TELUS Digital

Canada

Remote

CAD 90,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading digital consultancy in Canada is seeking a Site Reliability Engineer (SRE) to support mission-critical systems handling sensitive patient data. This role involves architecting and maintaining cloud infrastructure across AWS and GCP, ensuring high availability and security. Ideal candidates will have strong experience with Terraform, Python, and cloud services, alongside a background in healthcare environments. The position offers the flexibility of remote work.

Qualifications

  • Experience deploying and maintaining infrastructure across GCP and AWS.
  • Proficiency in Infrastructure-as-Code methodologies using Terraform.
  • Strong understanding of networking and distributed systems.
  • Knowledge of healthcare compliance and reliability standards.

Responsibilities

  • Architect and maintain secure, scalable cloud infrastructure.
  • Build automation tools using Python for operational excellence.
  • Design and implement monitoring and observability systems.
  • Collaborate with engineering and clinical operations teams.

Skills

Hands-on experience with GCP
Strong proficiency with Terraform
Solid understanding of Linux systems
Experience with GitHub Actions
Working knowledge of Python
Familiarity with healthcare data environments
Experience with JIRA and Confluence
Understanding of SRE principles

Tools

Terraform
GitHub Actions
Docker
Kubernetes

Job description

Welcome to TELUS Digital — where innovation drives impact at a global scale. As an award-winning digital product consultancy and the digital division of TELUS, one of Canada’s largest telecommunications providers, we design and deliver transformative customer experiences through cutting‑edge technology, agile thinking, and a people‑first culture.

With a global team across North America, South America, Central America, Europe, and APAC, we offer end‑to‑end expertise across eight core service areas: Digital Product Consulting, Digital Marketing Services, Data & AI, Strategy Consulting, Business Operations Modernization, Enterprise Applications, Cloud Engineering, and QA & Test Engineering.

From mobile apps and websites to voice UI, chatbots, AI, customer service, and in‑store solutions, TELUS Digital enables seamless, trusted, and digitally powered experiences that meet customers wherever they are — all backed by the secure infrastructure and scale of our multi‑billion‑dollar parent company.

Location & Flexibility

This role will work from home (Canada).

Opportunity: Site Reliability Engineer (SRE) – EMR
About the Role

We are seeking a Site Reliability Engineer (SRE) with strong cloud, automation, and infrastructure expertise to support our mission‑critical EMR. In this role, you will ensure the reliability, security, and performance of systems that handle sensitive patient information and clinical workflows. You’ll work across GCP, AWS, and modern DevOps tooling to build resilient infrastructure that meets healthcare‑grade compliance and uptime requirements.

Key Responsibilities
Cloud Infrastructure & Platform Reliability
  • Architect, deploy, and maintain secure, scalable infrastructure across GCP and AWS for EMR/medical data workloads.
  • Implement and manage Terraform‑based Infrastructure‑as‑Code to support consistent, compliant environment provisioning.
  • Ensure systems meet healthcare reliability standards, including high availability, disaster recovery, and data durability.
Automation & Operational Excellence
  • Build automation tools and scripts using Python to reduce manual operations and improve system consistency.
  • Enhance release deployment process using GitHub Actions to support safe, traceable, and compliant deployments.
  • Develop runbooks, operational workflows, and automated remediation for common reliability issues.
Monitoring, Observability & Incident Response
  • Design and maintain monitoring, alerting, and observability systems for EMR/clinical applications.
  • Lead incident response, root‑cause analysis, and post‑incident reviews with a focus on long‑term reliability improvements.
  • Define and track SRE metrics such as SLIs, SLOs, and error budgets tailored to clinical system uptime requirements.
Collaboration & Process Management
  • Work closely with engineering, data, and clinical operations teams to design reliable architectures for medical data systems.
  • Use JIRA to manage tasks, incidents, and sprint workflows.
  • Document infrastructure, processes, and compliance artifacts in Confluence.
Required Skills & Experience
  • Hands‑on experience with GCP and AWS cloud platforms.
  • Strong proficiency with Terraform and Infrastructure‑as‑Code methodologies.
  • Solid understanding of Linux systems, networking, and distributed systems.
  • Experience with GitHub Actions for version control and deployment automation.
  • Working knowledge of Python for scripting and automation.
  • Familiarity with healthcare data environments, EMR/EHR systems, or regulated data workflows.
  • Experience with JIRA and Confluence in an Agile environment.
  • Understanding of SRE principles: SLIs, SLOs, error budgets, incident management.
Nice‑to‑Have
  • Experience with containerization (Docker, Kubernetes).
  • Knowledge of healthcare interoperability standards (HL7, FHIR).
  • Experience with monitoring tools (Cloud Monitoring, CloudWatch).
  • Background in security engineering or compliance frameworks.
What You’ll Bring

You’re someone who thrives in high impact environments where reliability truly matters. You care deeply about automation, secure design, and building systems that clinicians and patients can depend on. You’re collaborative, curious, and committed to operational excellence.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

TEEMA • Vancouver

On-site
CAD 83,000 - 110,000
Site Reliability Engineer
Site Reliability Engineer

Tecsys Inc. • Montreal (administrative region)

On-site
CAD 90,000 - 120,000
Digital-first work environment
Collaborative workspaces
Continuous learning opportunities
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

Veerum • Calgary

Hybrid
CAD 115,000 - 130,000
Flexible benefits
Strong time-off policies
Continuous learning opportunities
Site Reliability Engineer
Site Reliability Engineer

Compunnel, Inc. • Montreal (administrative region)

Hybrid
CAD 90,000 - 130,000
Staff Site Reliability Developer, Protected Data SRE
Staff Site Reliability Developer, Protected Data SRE

Socket.dev • Southwestern Ontario

On-site
CAD 216,000 - 221,000
Site Reliability Engineer (SRE) – Observability
Site Reliability Engineer (SRE) – Observability

Astra-North Infoteck Inc. ~ Conquering today’s challenges, achieving tomorrow’s vision! • Toronto

Hybrid
CAD 75,000 - 95,000
Staff Site Reliability Developer, Google Unified Security and Threat Operations
Staff Site Reliability Developer, Google Unified Security and Threat Operations

Google • Southwestern Ontario

On-site
CAD 216,000 - 221,000
Staff Site Reliability Developer, Protected Data SRE
Staff Site Reliability Developer, Protected Data SRE

Google • Southwestern Ontario

On-site
CAD 216,000 - 221,000
Bonus
Equity
Benefits
Site Reliability Manager, Data center Networking, SRE
Site Reliability Manager, Data center Networking, SRE

Google • Southwestern Ontario

On-site
CAD 216,000 - 221,000
Equity
Bonus target
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Sage Recruiting Inc. • Canada

On-site
CAD 180,000 - 200,000
Unlimited vacation
Comprehensive health and dental benefits