Site Reliability Engineer I

Pager

Toronto

Hybrid

CAD 136,000 - 206,000

Full time

4 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Company equity
ESPP
Retirement plan
Paid vacation
Paid holidays
Mental wellness programs
Volunteer time off

Job summary

PagerDuty, Inc. is seeking an early-career Site Reliability Engineer I on the Core Infrastructure team.

You will help build and operate the foundation that powers PagerDuty's real-time digital operations platform, supporting millions of events and alerts daily while collaborating across teams to improve reliability and performance. The role requires Linux production experience, Kubernetes/EKS, cloud familiarity (AWS/GCP/Azure), and IaC skills.

Qualifications

  • 0 to 1+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles.
  • Hands‑on experience operating Linux-based systems in production environments.
  • Working knowledge of networking fundamentals, such as load balancing, DNS, TLS, and ingress traffic flow.
  • Experience with container orchestration (e.g., EKS, Kubernetes).
  • Experience working on cloud-native infrastructure (e.g., AWS, GCP, Azure), including networking and compute concepts.
  • Proficiency in at least one programming language (e.g., Python, Ruby, Go, etc.).
  • Experience with Infrastructure as Code (e.g., Terraform, CloudFormation).

Responsibilities

  • Support and improve foundational infrastructure, including networking, compute, Kubernetes clusters, and ingress/traffic management systems.
  • Contribute to the reliability and scalability of PagerDuty's core platform by hardening existing systems and supporting the rollout of new infrastructure capabilities.
  • Participate in agile rituals (standups, planning, retros) and communicate progress/risks early
  • Stay current on technical trends to suggest innovative tools and approaches to interesting problems
  • Monitor system health using metrics, logs, and alerts, and participate in 24/7 on-call rotations to help detect, respond to, and resolve incidents.

Skills

Linux systems
Networking basics
Kubernetes
Cloud platforms
Python/Ruby/Go
Terraform/CloudFormation

Tools

Kubernetes (EKS)
AWS
GCP
Azure
DataDog
New Relic
Prometheus/Grafana
Envoy/Istio/NGINX

Job description

PagerDuty, Inc. (NYSE: PD) is the global leader in AI-first digital operations. By automatically detecting, diagnosing, and remediating issues, the PagerDuty Platform orchestrates AI agents and automated workflows with context from over 750 integrations. Trusted by approximately two-thirds of the Fortune 100 and nearly half of the Fortune 500, PagerDuty is the industry standard for organizations scaling resilient, autonomous operations. Notable customers include Chipotle, Cloudflare, Docusign, Fox, Nvidia, Salesforce, Spotify, Zoom and more. We are growing rapidly and hiring top talent with leading AI skills across engineering, sales, product, marketing, and beyond as we build the leading digital operations platform.

As a Site Reliability Engineer I on the Core Infrastructure team in our Atlanta office, you'll help build and operate the foundational infrastructure that powers PagerDuty's real-time digital operations platform. Our systems support millions of events and alerts daily, enabling customers to detect, respond to, and resolve incidents quickly and reliably. You'll work at the intersection of platform evolution and operational excellence, building and evolving foundational network, compute, and ingress infrastructure while scaling and hardening existing systems. Your work will directly impact the reliability, scalability, and security of the services our customers rely on to keep their businesses running as PagerDuty continues to grow across products, regions, and customer use cases.

Key Responsibilities
  • Support and improve foundational infrastructure, including networking, compute, Kubernetes clusters, and ingress/traffic management systems.
  • Contribute to the reliability and scalability of PagerDuty's core platform by hardening existing systems and supporting the rollout of new infrastructure capabilities.
  • Participate in agile rituals (standups, planning, retros) and communicate progress/risks early
  • You stay current on technical trends to suggest innovative tools and approaches to interesting problems
  • Monitor system health using metrics, logs, and alerts, and participate in 24/7 on-call rotations to help detect, respond to, and resolve incidents.
Basic Qualifications
  • 0 to 1+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles
  • Hands‑on experience operating Linux-based systems in production environments
  • Working knowledge of networking fundamentals, such as load balancing, DNS, TLS, and ingress traffic flow
  • Experience with container orchestration (e.g., EKS, Kubernetes)
  • Experience working on cloud-native infrastructure (e.g., AWS, GCP, Azure), including networking and compute concepts
  • Proficiency in at least one programming language (e.g., Python, Ruby, Go, etc.)
  • Experience with Infrastructure as Code (e.g., Terraform, CloudFormation)
Preferred Qualifications
  • Experience with AWS cloud networking concepts such as VPCs, subnets, routing, security groups, and load balancers
  • Experience operating or contributing to production Kubernetes platforms (e.g., EKS), including cluster upgrades, networking, or ingress configuration
  • Experience with monitoring, observability, and logging platforms (e.g., DataDog, New Relic, SumoLogic, Splunk, Prometheus, Grafana)
  • Familiarity with service meshes, ingress controllers, or API gateways (e.g., Envoy, Istio, NGINX)

Salary Range: $98,000 to $148,500

Where we work

PagerDuty operates a hybrid work model with offices in 8 major cities: Atlanta, Lisbon, London, San Francisco, Santiago, Sydney, Tokyo, and Toronto. While we offer flexibility within our established locations, we cannot employ candidates residing in:

Location restrictions: Australia: Northern Territory, Queensland, South Australia, Tasmania, Western Australia
Canada: Alberta, Manitoba, Newfoundland, Northwest Territories, Nunavut, PEI, Quebec, Saskatchewan, Yukon
United States: Alaska, Hawaii, Iowa, Louisiana, Mississippi, Nebraska, New Mexico, Oklahoma, Rhode Island, South Dakota, West Virginia, Wyoming

Candidates must reside in an eligible location, which vary by role.

How we work

Our values guide how we support customers, collaborate with colleagues, develop products, and foster a culture of belonging. They define not just our actions, but what it means to be Dutonian.

People Leaders at PagerDuty are responsible for creating high performance environments that drive accountability. PagerDuty has four key dimensions that define our Leadership Impact: Lead Self, Lead the Team, Lead the Business, and Lead the Future. Each dimension has three associated competencies to give leaders a shared language for guiding their development, career, promotion, and succession planning discussions. Our Manager Expectations serve as a practical guide for managers to understand their responsibilities, prioritize their efforts, and drive engagement and performance.

What we offer

As a global organization, our total rewards approach is competitive with industry standards and aligned with local laws and regulations. Learn more, including country-specific offerings, on our benefits site .

Your package may include:
  • Company equity*
  • ESPP (Employee Stock Purchase Program)*
  • Retirement or pension plan*
  • Generous paid vacation time
  • Paid holidays and sick leave
  • Dutonian Wellness Days & HibernationDuty - companywide paid days off in addition to PTO
  • Paid parental leave: 22 weeks for pregnant parent, 12 weeks for non-pregnant parent (some countries have longer leave standards and we comply with local laws)*
  • Paid volunteer time off: 20 hours per year
  • Mental wellness programs

*Eligibility may vary by role, region, and tenure

About PagerDuty

PagerDuty, Inc. (NYSE:PD) is a global leader in digital operations management. The PagerDuty Operations Cloud is an AI-powered platform that empowers business resilience and drives operational efficiency for enterprises. With a generative AI assistant at its core, PagerDuty empowers teams to detect and resolve issues in real time, orchestrate complex workflows, and drive continuous improvement across their digital operations. Trusted by nearly half of both the Fortune 500 and the Forbes AI 50, as well as approximately two-thirds of the Fortune 100, PagerDuty is essential for delivering always-on digital experiences to modern businesses

PagerDuty is Great Place to Work-certified, a Fortune Best Workplace for Millennials, a Fortune Best Medium Workplace, a Fortune Best Workplace in Technology, and a top rated product on TrustRadius and G2.

Go behind-the-scenes on our careers site and @pagerduty on Instagram.

Additional Information

PagerDuty is an equal opportunity employer. PagerDuty does not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, parental status, veteran status, or disability status. Your privacy is important to us. By submitting an application, you confirm that you have read and understand PagerDuty's Privacy Policy .

PagerDuty is committed to providing reasonable accommodations for qualified individuals with disabilities in our job application process. Should you require accommodation, please email accommodation@pagerduty.com and we will work with you to meet your accessibility needs.

PagerDuty uses the E-Verify employment verification program.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Solutions Consultant New Toronto
Senior Solutions Consultant New Toronto

Pager • Toronto

On-site
CAD 144,000 - 197,000
Company equity
ESPP (Employee Stock Purchase Program)
Retirement or pension plan
+6
Senior Solutions Consultant
Senior Solutions Consultant

PagerDuty, Inc. • Toronto

Hybrid
CAD 144,000 - 197,000
Competitive salary
Comprehensive benefits package
Equity
+1
Senior Solutions Consultant
Senior Solutions Consultant

Triwill Group • Toronto

Hybrid
CAD 144,000 - 197,000
Competitive salary
Comprehensive benefits package
Flexible work arrangements
+5
Senior Solutions Consultant
Senior Solutions Consultant

PagerDuty • Toronto

Hybrid
CAD 144,000 - 197,000
Flexible work arrangements
Comprehensive benefits package
Company equity
+5
Senior Principal Industry Analyst Relations
Senior Principal Industry Analyst Relations

PagerDuty • Toronto

Hybrid
CAD 153,000 - 231,000
Competitive salary
Comprehensive benefits package
Flexible work arrangements
+7
Senior Principal Industry Analyst Relations
Senior Principal Industry Analyst Relations

Triwill Group • Toronto

Hybrid
CAD 153,000 - 231,000
Comprehensive benefits
Flexible work arrangements
Equity participation
+1
Senior Principal Industry Analyst Relations
Senior Principal Industry Analyst Relations

Pager • Toronto

Hybrid
CAD 153,000 - 231,000
Company equity
ESPP
Retirement plan
+5
Senior Compensation Partner
Senior Compensation Partner

PagerDuty • Toronto

Hybrid
CAD 98,000 - 149,000
Competitive salary
Comprehensive benefits package
Flexible work arrangements
+2
Account Manager - Toronto
Account Manager - Toronto

Triwill Group • Toronto

On-site
CAD 105,000 - 127,000
Competitive salary
Comprehensive benefits
Flexible work arrangements
+7
Account Manager - Toronto
Account Manager - Toronto

PagerDuty • Toronto

Hybrid
CAD 105,000 - 127,000
Competitive salary
Comprehensive benefits
Flexible work arrangements
+4