Intermediate Site Reliability Engineer New Toronto

ContactMonkey Inc.

Toronto

Hybrid

CAD 130,000 - 150,000

Full time

14 hours ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Employer-paid benefits
Work from anywhere up to 4 weeks
Stock option plan
RRSP group savings plan
Generous vacation
Personal development budget
Personal day
Volunteer days
Birthday off
Health days

Job summary

ContactMonkey Inc. is seeking an Intermediate Site Reliability Engineer to own reliability of our AWS and Kubernetes-based platforms.

You will collaborate with SRE and development teams to enhance deployments, investigate issues, and address security risks, including SOC 2 and GDPR controls. You’ll work with Terraform, Terragrunt, Kubernetes, and Docker, contribute to CI/CD improvements, and participate in on-call rotations across multi-region AWS environments.

Qualifications

  • 3–5 years in SRE/DevOps or related role.
  • Terraform modules & Terragrunt configuration experience.
  • Docker & Kubernetes troubleshooting and deployments.
  • Solid Linux, networking, DNS, HTTP, TLS understanding.
  • Git, CI/CD pipelines, and deployment workflows.
  • Experience using logs, metrics, dashboards, and alerts to investigate production problems.
  • Application security experience and vulnerability remediation.

Responsibilities

  • Maintain AWS and Kubernetes environments and improve availability, performance, and resource usage.
  • Build and maintain infrastructure as code with Terraform/Terragrunt; manage environment configurations.
  • Improve CI/CD pipelines, deployment automation, and rollback procedures.
  • Enhance monitoring, incident response, and on-call support.
  • Collaborate with developers to address security risks and SOC 2/GDPR controls.
  • Strengthen IAM, secrets management, and network security in cloud environments.

Skills

SRE/DevOps
Cloud operations
Linux networking
Security tooling
Incident response
Communication

Tools

Terraform
Terragrunt
Kubernetes
Docker
GitHub Actions
Prometheus
Grafana
CloudWatch

Job description

Our mission is to power measurable employee engagement worldwide, and we're looking for an Intermediate Site Reliability Engineer to join our Engineering team.

About the job

We're looking for someone who enjoys running production systems, understands how applications work, and brings practical application security experience.

You'll work closely with our SRE and development teams to maintain our infrastructure, improve deployments, investigate production issues, and address security risks. You'll also contribute to the technical controls that support SOC 2 audits and GDPR compliance.

Our environment includes AWS, Kubernetes on EKS, Terraform, Terragrunt, GitHub Actions, Prometheus, Grafana, and CloudWatch. Our applications use Ruby on Rails, Vue.js, and Node.js, with MySQL, PostgreSQL, and Sidekiq supporting the backend.

You'll take ownership of defined projects and operational improvements, with senior engineers available for guidance and review. There's room to develop deeper expertise in reliability, infrastructure automation, and application security.

Your impact

  • Infrastructure & reliability: Maintain AWS and Kubernetes environments, troubleshoot production issues, and improve availability, performance, and resource usage.
  • Terraform & Terragrunt: Build and maintain infrastructure as code, review plans, manage environment configuration, and address infrastructure drift.
  • Deployments & developer experience: Improve CI/CD pipelines, deployment automation, release checks, and rollback procedures.
  • Monitoring & incidents: Improve monitoring and alerts, join the on-call rotation, and contribute to incident reviews.
  • Application security: Work with developers to assess vulnerabilities, review security risks, and validate fixes.
  • CI/CD security: Maintain code, dependency, secrets, container, and infrastructure scanning.
  • Cloud security: Strengthen IAM, secrets management, network controls, and Kubernetes security.
  • SOC 2 & GDPR: Support technical controls, audit evidence, and personal data protection.
  • Recovery: Test backups and recovery procedures, maintain runbooks, and support production readiness.
  • Collaboration: Participate in code reviews, document changes, and support application, AI, and data engineering teams.

About you

  • Around 3–5 years of experience in SRE, DevOps, platform engineering, cloud operations, or a related engineering role. Equivalent practical experience is welcome.
  • Experience writing and maintaining Terraform modules and Terragrunt configuration, including reviewing plans and working with remote state.
  • Experience with Docker and Kubernetes, including troubleshooting deployments, services, health checks, and resource limits.
  • A solid understanding of Linux, networking, DNS, HTTP, and TLS.
  • Experience with Git, pull requests, CI/CD pipelines, and deployment workflows.
  • Experience using logs, metrics, dashboards, and alerts to investigate production problems.
  • Practical application security experience through vulnerability remediation, secure code review, threat modelling, or security tooling.
  • Understanding of common web application risks, including broken access control, injection, authentication weaknesses, and sensitive data exposure.
  • Familiarity with IAM, least privilege, secrets management, and encryption.
  • Working knowledge of SOC 2 controls and GDPR principles relevant to engineering, including access restrictions, data minimization, retention, and deletion.
  • Clear communication skills and good judgment about when to work independently, request a review, or elevate an issue.

How you can stand out

  • Supported Ruby on Rails or Node.js applications in production
  • Worked with MySQL, PostgreSQL, Redis or Valkey, and Sidekiq
  • Experience with GitHub Actions, Argo CD, Helm, Karpenter, or KEDA
  • Built useful monitoring with Prometheus, Grafana, or CloudWatch
  • Supported services across multiple AWS regions
  • Helped remediate penetration-test findings or contributed technical evidence to a SOC 2 audit
  • Participated in backup restoration or disaster recovery exercises.
  • Familiar with securing AI integrations, agent workloads, or MCP services
  • You hold relevant AWS, Kubernetes, Terraform, or security certifications.

Working with the team


You'll work closely with SRE and application engineers, and support AI and data initiatives where they depend on shared infrastructure.

We value people who ask questions, explain their reasoning, and leave systems easier for others to understand. You'll take part in technical discussions and code reviews, share what you learn, and help improve how we operate.

Reducing recurring incidents, unnecessary alerts, and repetitive manual work is part of the role. We want reliable systems and sustainable operations for the people supporting them.

What we bring to the table
  • 100% employer-paid benefits and a Health Spending Account from day one
  • Work from anywhere in the world for up to four weeks
  • A stock option plan so you can own a piece of our success
  • An RRSP Group Savings Plan
  • A generous vacation package
  • A personal development budget
  • One personal day and two volunteering days
  • Your birthday off
  • Five health days per year
  • A downtown Toronto office for hybrid work, with plenty of snacks

Compensation and work details

The salary range for this role is $130,000-$150,000 Compensation is based on experience, skills, and our internal compensation framework and equity.

We're happy to discuss compensation throughout the hiring process.

This is a full-time position on our SRE team.

The role includes a shared on-call rotation after onboarding. We'll discuss the schedule, escalation support, and expectations during the interview process.

Who we are


ContactMonkey helps organizations create, send, and measure internal communications directly within Outlook and Gmail.

Our platform brings together email design, employee engagement tools, and analytics so internal communications teams can understand what reaches their people and what gets a response.

As the product grows, we're investing in reliability, security, and tooling that helps our engineering teams deliver changes confidently.

Diversity is our strength
At ContactMonkey, we’re building products for diverse organizations, and we need a diverse team to do that. We strongly encourage applications from everyone regardless of race, religion, colour, national origin, gender, sexual orientation, age, marital status, or disability status.

We are committed to an accessible hiring process. If you need accommodations or adjustments during interviews or beyond, please let us know so we can arrange the support you need.

AI Disclosure
We use AI to take notes during our interviews. Applications and interviews are reviewed by our Talent Acquisition team. Our applicant tracking system uses AI for workflows and hiring process efficiencies.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Digital Customer Success Manager, SMB
Digital Customer Success Manager, SMB

ContactMonkey Inc. • Toronto

Hybrid
CAD 80,000 - 120,000
Health Spending Account
Work from anywhere for up to 4 weeks
Stock option plan
+7
Applied AI Engineer
Applied AI Engineer

Communitech • Toronto

On-site
CAD 140,000 - 160,000
100% employer-paid benefits
Stock option plan
Downtown Toronto office with hybrid in
Sales Development Representative
Sales Development Representative

ContactMonkey Inc. • Toronto

Hybrid
CAD 55,000 - 80,000
Employer-paid benefits and HSA
Work from anywhere up to 4 weeks
Stock option plan
+7
Strategic Account Executive New Toronto
Strategic Account Executive New Toronto

ContactMonkey Inc. • Toronto

On-site
CAD 225,000 - 250,000
Employer-paid benefits
Stock options
RRSP
+3
Strategic Account Executive
Strategic Account Executive

ContactMonkey • Toronto

Hybrid
CAD 225,000 - 250,000
Health benefits
Stock option plan
RRSP Group Savings Plan
+4
Campaigns Manager New Toronto
Campaigns Manager New Toronto

ContactMonkey Inc. • Toronto

Hybrid
CAD 120,000 - 140,000
Employer-paid benefits
Work from anywhere up to 6 weeks
Stock option plan
+2
Campaigns Manager
Campaigns Manager

ContactMonkey Inc. • Toronto

Hybrid
CAD 120,000 - 150,000
Employer-paid benefits
Stock options
Generous vacation
+6
Applied AI Engineer New Toronto
Applied AI Engineer New Toronto

ContactMonkey Inc. • Toronto

On-site
CAD 140,000 - 160,000
100% employer-paid benefits
Work from anywhere in the world up to
Stock option plan
+5
New Business Account Executive
New Business Account Executive

SurveyMonkey • Ottawa

Hybrid
CAD 56,000 - 66,000
Medical Insurance
Dental Insurance
Vision Insurance
+8
Software Engineering Team Lead
Software Engineering Team Lead

Monks • Toronto

On-site
CAD 130,000 - 160,000