Platform & SRE Engineer

TechDoQuest

Montreal (administrative region)

On-site

CAD 90,000 - 130,000

Full time

41 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

TechDoQuest is seeking a DevOps/Platform Engineer to design, deploy, and manage scalable Kubernetes-based services in a fast-paced environment. You will build and maintain CI/CD pipelines, strengthen observability with Prometheus and Grafana, and ensure secure service onboarding with OAuth2/SSO.

You will troubleshoot issues, automate tasks, and support API and UI deployments. Ideal candidates are hands-on with Linux, container orchestration, and web deployment architectures, and collaborate

Qualifications

  • Kubernetes and Docker hands-on experience.
  • Solid understanding of containerization, deployment strategies, and orchestration.
  • Deep familiarity with observability stacks: Prometheus, Grafana, alerts, and metrics.
  • Working knowledge of web deployment and web component architecture.
  • Experience with authN/authZ mechanisms (OAuth, SSO).
  • Good understanding of CI/CD pipelines (Jenkins, GitLab CI).
  • Comfort with Linux and basic database operations (SQL).

Responsibilities

  • Design, deploy, and manage scalable, containerized applications on Kubernetes clusters.
  • Build and maintain CI/CD pipelines for reliable and repeatable deployments.
  • Own and improve platform observability using Prometheus and Grafana with effective alerts.
  • Collaborate with application teams to onboard services with proper authentication/authorization.
  • Troubleshoot production issues, improve availability, and drive root cause analysis.
  • Automate routine tasks using Python scripting, Ansible, or similar tooling.
  • Support web-based deployments including APIs and UI apps.
  • Collaborate on system design, capacity planning, and performance tuning.

Skills

Kubernetes
Docker
Containerization
Observability
OAuth/SSO
CI/CD
Linux
SQL

Tools

Prometheus
Grafana
Ansible
Terraform
Helm

Job description


  • Design, deploy, and manage scalable, containerized applications on Kubernetes clusters.

  • Build and maintain CI/CD pipelines to enable reliable and repeatable deployments.

  • Own and improve platform observability using tools like Prometheus, Grafana, and alerts.

  • Work closely with application teams to onboard services with proper authentication and authorization (OAuth2, SSO, JWT, etc.).

  • Troubleshoot production issues, improve availability, and drive root cause analysis.

  • Standardize and automate routine tasks using Python scripting or Ansible.

  • Support web-based deployments including APIs, UI apps, and their configurations.

  • Collaborate on system design, capacity planning, and performance tuning.


Must-Have Skills


  • Solid hands-on experience with Kubernetes and Docker.

  • Strong understanding of containerization, deployment strategies, and orchestration.

  • Deep familiarity with observability stacks: Prometheus, Grafana, alerts, and metrics.

  • Working knowledge of web deployment and web component architecture.

  • Experience in setting up or working with authN/authZ mechanisms (OAuth, SSO)

  • Good understanding of CI/CD pipelines (Jenkins, GitLab CI, etc.).

  • Comfort with Linux and basic database operations (SQL)


Nice to Have


  • Proficiency in Python scripting for automation.

  • Experience with Ansible or other config management tools.

  • Exposure to infrastructure as code (Terraform, Helm).

  • Familiarity with cloud-native concepts and services.


Who You Are


  • A systems thinker who understands both the development and operations side.

  • Someone who thrives in a fast-paced, cross-functional team.

  • Passionate about automation, stability, and continuous improvement.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Mantu • Montreal (administrative region)

On-site
CAD 90,000 - 130,000
Apps Development Sr Manager - Vice President
Apps Development Sr Manager - Vice President

Citigroup Inc. • Mississauga

On-site
CAD 100,000 - 130,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

twentysix • Vancouver

On-site
CAD 90,000 - 130,000
Senior DevOps Engineer
Senior DevOps Engineer

MarkiTech • Toronto

On-site
CAD 120,000 - 180,000
Senior Site Reliability Engineer, SRE
Senior Site Reliability Engineer, SRE

Jobtailor • Toronto

On-site
CAD 120,000 - 180,000
[8SN] Senior Site Reliability Engineer (SRE) – Kubernetes
[8SN] Senior Site Reliability Engineer (SRE) – Kubernetes

Worky • Montreal (administrative region)

On-site
CAD 120,000 - 170,000
Laptop
Flexible work arrangements
Professional development and training
Infrastructure Developer – Platform Engineering
Infrastructure Developer – Platform Engineering

Jobtailor • Montreal (administrative region)

On-site
CAD 90,000 - 130,000
Senior DevOps Engineer
Senior DevOps Engineer

Mphasis • Toronto

On-site
CAD 120,000 - 180,000
DevOps Specialist - Kubernetes
DevOps Specialist - Kubernetes

Myticas Consulting • Ottawa

On-site
CAD 90,000 - 130,000
Senior DevOps Engineer
Senior DevOps Engineer

MarkiTech.AI • Toronto

On-site
CAD 140,000 - 180,000