Site Reliability Engineer (OpenShift & Infrastructure)

Accion Labs

Canada

On-site

CAD 90,000 - 115,000

Part time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A tech company in Canada is seeking a Site Reliability Engineer with expertise in OpenShift and automated provisioning. The role involves managing clusters, ensuring high availability, and leading incident responses. Ideal candidates will be experienced in SRE practices and knowledgeable in using Terraform and Ansible. This position is contract-based and targets mid-senior level professionals.

Qualifications

  • Experience with managing OpenShift clusters in cloud and on-premise.
  • Skill in automating platform provisioning using tools like Terraform and Ansible.
  • Proficiency in configuring monitoring tools like Prometheus and Grafana.

Responsibilities

  • Install, configure, and upgrade OpenShift clusters.
  • Manage internal networking and cluster services.
  • Lead incident response and root cause analysis.

Skills

OpenShift administration
GitOps workflows
Terraform
Ansible
TLS/MTLS encryption
Monitoring and alerting frameworks
Incident response
SRE best practices
LDAP authentication
Networking and DNS management

Job description

Site Reliability Engineer (OpenShift & Infrastructure)

Get AI-powered advice on this job and more exclusive features.

Direct message the job poster from Accion Labs

Responsibilities & Skills
  • Install, configure, upgrade, and administer OpenShift clusters (OCP) in on-premise and cloud environments.
  • Manage OCP internal networking, ingress, egress, and cluster services.
  • Configure and integrate LDAP authentication and access management.
  • Implement TLS and MTLS encryption, and manage certificate lifecycle for secure communications.
  • Implement GitOps workflows using ArgoCD for continuous delivery and environment consistency.
  • Automate platform and application provisioning using Terraform and Ansible.
  • Configure and maintain F5 LTM load balancers.
  • Configure and manage DNS, networking, and subnets.
  • Build and manage monitoring, logging, and alerting frameworks (e.g., Prometheus, Grafana, ELK).
  • Define and enforce SLIs/SLOs and error budgets for services running on OCP.
  • Lead incident response, root cause analysis (RCA), and postmortems.
  • Build automation for self‑healing, scaling, and zero-touch operations.
  • Ensure high availability, disaster recovery, and failover strategies are implemented.
  • Secure platform and workloads following enterprise security standards.
  • Support application deployments and CI/CD pipelines on OpenShift.
  • Troubleshoot networking, cluster, and deployment issues end-to-end.
  • Apply SRE best practices to improve reliability, scalability, and performance.
  • Collaborate with development and platform teams to optimize system operations.
Seniority level
  • Mid‑Senior level
Employment type
  • Contract
Job function
  • Information Technology
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Apps Development Sr Manager - Vice President
Apps Development Sr Manager - Vice President

Citi • Mississauga

On-site
CAD 100,000 - 130,000
Apps Development Sr Manager - Vice President
Apps Development Sr Manager - Vice President

Citigroup Inc. • Mississauga

On-site
CAD 100,000 - 130,000
Site Reliability Engineer (Linux / Cloud Infrastructure)
Site Reliability Engineer (Linux / Cloud Infrastructure)

Atlantis IT Group • Montreal

On-site
CAD 80,000 - 100,000
Senior DevOps Engineer (SRE Focus) – Kubernetes | OpenShift | CI/CD |
Senior DevOps Engineer (SRE Focus) – Kubernetes | OpenShift | CI/CD |

United States Digital Space LLC • Mississauga

On-site
CAD 120,000 - 160,000
Operation Support Analyst
Operation Support Analyst

Kumaran Systems • Toronto

On-site
CAD 70,000 - 100,000
[8SN] Senior Site Reliability Engineer (SRE) – Kubernetes
[8SN] Senior Site Reliability Engineer (SRE) – Kubernetes

Worky • Montreal (administrative region)

On-site
CAD 120,000 - 170,000
Laptop
Flexible work arrangements
Professional development and training
DevOps Consultant - BuildMaster & Openshift
DevOps Consultant - BuildMaster & Openshift

Bevertec • Alberta

Remote
CAD 110,000 - 160,000
Senior Site Reliability Engineer (SRE) – Kubernetes
Senior Site Reliability Engineer (SRE) – Kubernetes

Software Mind • Montreal (administrative region)

Hybrid
CAD 120,000 - 170,000
Competitive salary
Laptop provided
Professional development
+2
DevOps Engineer
DevOps Engineer

NTT DATA North America • Toronto

On-site
CAD 90,000 - 120,000
Site Reliability Engineer
Site Reliability Engineer

Vertex Elite LLC • Ottawa

On-site
CAD 83,000 - 124,000