Senior Platform Engineer

XYZ Venture Capital

Toronto

On-site

CAD 120,000 - 180,000

Full time

14 days+
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Competitive compensation
Equity
Health, dental, vision
3 weeks vacation + unlimited sick days
Home office stipend
AI tools access

Job summary

Rootly is seeking an experienced Site Reliability Engineer to help shape infrastructure that underpins incident response for some of the world’s most forward‑thinking teams. You’ll own CI/CD pipelines, observability tooling, monitoring systems, and incident response processes to improve reliability and velocity.

You’ll collaborate across engineering to design scalable systems, define SLOs, and advocate for reliability. Ruby and Go experience is a plus.

Qualifications

  • 5+ years of experience in an SRE/Platform/DevOps or Infrastructure Engineering role.
  • 5+ years of experience writing software in a production environment.
  • Strong knowledge of cloud infrastructure, distributed systems, and reliability practices.
  • Experience with observability, performance tuning, and scaling strategies.
  • Familiarity with incident response, monitoring, and CI/CD systems.
  • Hands-on experience supporting web services at meaningful scale.

Responsibilities

  • Embed with product teams to enhance observability, reliability, and performance of services.
  • Own CI/CD pipelines, observability tooling, monitoring systems, and incident response processes.
  • Build tools and automation to reduce toil, improve velocity and reliability.
  • Collaborate across engineering to understand systems and surface reliability and scaling concerns early.
  • Architect and scale infrastructure for performance and availability.
  • Drive capacity planning to ensure resilience as we grow.
  • Define and manage SLOs and error budgets with production teams.
  • Advocate for reliability and performance across the engineering org.

Skills

SRE experience
Platform engineering
DevOps
Production software
Observability

Tools

Ruby
Go

Job description

About Rootly

At Rootly, we are on a mission to be the go-to way companies respond when things go wrong, helping every organization be more reliable. We do this by building an industry-leading incident management platform that allows companies around the world consistently and quickly resolve incidents. We are not simply transforming an industry, we are carving an entirely new +$B segment ourselves and need incredible talent to achieve this ambitious goal together.

Customers love Rootly. Some of the fastest growing companies around the world such as NVIDIA, Figma, Canva, Tripadvisor, Squarespace and more rely on Rootly to power their critical incident management process. They obsess over our delightful enterprise-ready platform and unique partnership model. See why our customers have reviewed us 5 stars on G2.

Investors love Rootly. We are backed by some of the most respected funds in the world from Y Combinator to operators like the CTO of Dropbox and GitHub. We'd be happy to disclose our entire funding and profitability picture live during the interview. As a culture we relentlessly put transparency first. We conduct monthly financial reviews as a team so everyone has a pulse on the health of the business and publish what we are building in our weekly changelog.

About The Role

This is a rare opportunity to join Rootly as an early engineer and fundamentally shape our trajectory. You’ll work on infrastructure that underpins incident response and on-call for some of the most forward‑thinking teams in the world. This is not a traditional ops role – we’re looking for software engineers who love infrastructure, developer experience, think in systems, and are obsessed with building tools that make the entire engineering org more reliable, performant, and scalable.

We move with speed, taste, and impact. At Rootly, engineers are expected to take initiative and deliver high‑leverage work—whether it’s building automation to remove toil, improving our observability stack, or making our services more resilient. You’ll work closely with product engineers and customers alike to ensure we’re always one step ahead. If you thrive in environments where ownership is real, excellence is expected, and reliability is non‑negotiable, this is the place for you.

What You’ll Do
  • Embed with product teams to enhance observability, reliability, and performance of their services.
  • Own our CI/CD pipelines, observability tooling, monitoring systems, and incident response processes.
  • Build tools and automation to eliminate manual toil, improve engineering velocity and developer experience, and improve system reliability.
  • Collaborate deeply across engineering to understand systems at the code level and surface cross‑cutting reliability, performance, and scaling concerns early.
  • Architect and scale our infrastructure, ensuring best‑in‑class performance, availability, and operational excellence.
  • Drive capacity planning efforts to ensure our infrastructure is resilient and scalable as we grow.
  • Define and manage SLOs and error budgets in partnership with Engineering teams who own production services.
  • Act as a strong voice for reliability, performance, and scalability across the engineering organization.
What You'll Need
Minimum Qualifications

You don’t need a fancy degree or a resume full of logos. What matters is your ability to execute, influence, and inspire. If the following sounds like you, we want to talk:

  • 5+ years of experience in an SRE, Platform, DevOps, or Infrastructure Engineering role.
  • 5+ years of experience writing software in a production environment.
  • Strong technical knowledge of cloud infrastructure, distributed systems, and reliability practices.
  • Strong understanding of observability, performance tuning, and scaling strategies.
  • Deep familiarity with incident response, monitoring, and CI/CD systems.
  • Hands‑on experience supporting web or RPC services at meaningful scale.
  • You write code to solve infrastructure problems—not shell scripts alone, but production‑grade software.
Preferred Qualifications
  • You have a big‑picture systems mindset and a proactive approach to reliability.
  • You’ve embedded with product teams and influenced design and architecture decisions.
  • You’re comfortable taking ownership of complex problems—and seeing them through.
  • Experience with Ruby and Go is a plus.
Our Tech Stack
  • Backend: Ruby on Rails, PostgreSQL, Redis, Sidekiq, AWS Lambdas, Kafka, DynamoDB
  • Frontend: Turbo, Stimulus, ViewComponents
  • Infrastructure: AWS, Terraform (IaC)
Why Rootly?

We’re not just another startup. We’re building something category‑defining and want teammates who crave ownership, love solving hard problems, and thrive in a high‑bar, high‑impact environment.

  • Competitive compensation and early equity in a fast‑growing, venture‑backed company.
  • Comprehensive medical, dental, and vision coverage.
  • 3 weeks of vacation, plus unlimited sick and mental health days, and a company‑wide end‑of‑year shutdown to recharge.
  • $500 stipend for home office setup.
  • Unlimited token usage and access to AI tools
  • A fast‑moving, high‑impact environment where your leadership and ideas directly shape the future of the company.

If this sounds like the kind of challenge and opportunity you’re looking for, apply now and let’s build something great together.

Rootly is an equal opportunity employer. We aim to create an environment where every team member at Rootly feels like they belong so they can have a greater impact on our business and customers. We do not discriminate on the basis of race, religion, colour, national origin, gender, sexual orientation, age, marital status, veteran status, or disability status.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Head of Platform Engineering
Head of Platform Engineering

XYZ Venture Capital • Toronto

On-site
CAD 180,000 - 260,000
Equity available
Medical, dental, vision coverage
3 weeks vacation + sick days
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

XYZ Venture Capital • Toronto

Hybrid
CAD 120,000 - 180,000
Equity
Medical coverage
Dental coverage
+7
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Rootly • Toronto

On-site
CAD 120,000 - 180,000
Competitive compensation
Comprehensive medical coverage
3 weeks of vacation
+3
Engineering Manager
Engineering Manager

Rootly • Toronto

Remote
CAD 140,000 - 190,000
Equity
Medical, dental, vision
3 weeks vacation
+3
Senior Product Engineer
Senior Product Engineer

Rootly • Toronto

On-site
CAD 95,000 - 120,000
Competitive compensation
Medical, dental, and vision coverage
3 weeks of vacation plus unlimited sick days
+1
Solutions Engineer
Solutions Engineer

Rootly • Toronto

On-site
CAD 80,000 - 110,000
Competitive compensation
Comprehensive medical, dental, and vision coverage
3 weeks of vacation plus unlimited sick days
+2
Solutions Engineer
Solutions Engineer

XYZ Venture Capital • Toronto

On-site
CAD 100,000 - 150,000
Equity
Medical/Dental/Vision
Vacation + sick leave
+2
Senior Product Engineer
Senior Product Engineer

XYZ Venture Capital • Toronto

On-site
CAD 130,000 - 180,000
Competitive compensation
Equity
Medical, dental, and vision coverage
+2
Senior AI Engineer
Senior AI Engineer

XYZ Venture Capital • Toronto

On-site
CAD 120,000 - 180,000
Equity in a fast-growing company
Comprehensive medical/dental/vision
Three weeks of vacation + sick days
+2
Senior Security Engineer
Senior Security Engineer

Rootly • Toronto

On-site
CAD 100,000 - 130,000
Competitive compensation
Comprehensive medical, dental, and vision coverage
3 weeks of vacation plus unlimited sick and mental health days
+2