Site Reliability Engineer I

PagerDuty

Atlanta (GA)

Hybrid

USD 98,000 - 149,000

Full time

2 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Competitive salary
Comprehensive benefits package
Flexible work arrangements
Company equity
ESPP
Retirement or pension plan
Generous paid vacation time

Job summary

PagerDuty, Inc. in Atlanta is seeking a Site Reliability Engineer I to join the Core Infrastructure team. You will help build and operate the foundational infrastructure powering PagerDuty's real-time digital operations platform.

You will work across networking, compute, and ingress systems to scale reliability and security for millions of events and alerts daily, while contributing to growth across products and regions.

Qualifications

  • Hands-on experience operating Linux-based systems in production environments.
  • Working knowledge of networking fundamentals: load balancing, DNS, TLS, ingress traffic flow.
  • Experience with container orchestration (Kubernetes/EKS).
  • Experience with cloud-native infrastructure (AWS, GCP, Azure).
  • Proficiency in at least one programming language (Python, Ruby, Go).
  • Experience with Infrastructure as Code (Terraform, CloudFormation).

Responsibilities

  • Support and improve foundational infrastructure, including networking, compute platforms, Kubernetes clusters, and ingress/traffic management systems.
  • Contribute to reliability and scalability by hardening existing systems and rolling out new infrastructure capabilities.
  • Participate in agile rituals and communicate progress/risks early.
  • Stay current on technical trends to suggest innovative tools and approaches.
  • Monitor system health using metrics, logs, and alerts; participate in 24/7 on-call rotations.

Skills

Linux systems
Networking basics
Kubernetes
Python
Go
Terraform
CloudFormation
CI/CD

Tools

AWS
GCP
Azure
DataDog
New Relic
Prometheus
Grafana
Envoy
NGINX
Terraform
Kubernetes

Job description

PagerDuty, Inc. (NYSE: PD) is the global leader in AI-first digital operations. By automatically detecting, diagnosing, and remediating issues, the PagerDuty Platform orchestrates AI agents and automated workflows with context from over 750 integrations. Trusted by approximately two-thirds of the Fortune 100 and nearly half of the Fortune 500, PagerDuty is the industry standard for organizations scaling resilient, autonomous operations. Notable customers include Chipotle, Cloudflare, Docusign, Fox, Nvidia, Salesforce, Spotify, Zoom and more. We are growing rapidly and hiring top talent with leading AI skills across engineering, sales, product, marketing, and beyond as we build the leading digital operations platform.

As a Site Reliability Engineer I on the Core Infrastructure team in our Atlanta office, you'll help build and operate the foundational infrastructure that powers PagerDuty's real‑time digital operations platform.

Our systems support millions of events and alerts daily, enabling customers to detect, respond to, and resolve incidents quickly and reliably.

You’ll work at the intersection of platform evolution and operational excellence, building and evolving foundational network, compute, and ingress infrastructure while scaling and hardening existing systems. Your work will directly impact the reliability, scalability, and security of the services our customers rely on to keep their businesses running as PagerDuty continues to grow across products, regions, and customer use cases.

Key Responsibilities
  • Support and improve foundational infrastructure, including networking, compute platforms, Kubernetes clusters, and ingress/traffic management systems.
  • Contribute to the reliability and scalability of PagerDuty’s core platform by hardening existing systems and supporting the rollout of new infrastructure capabilities.
  • Participate in agile rituals (standups, planning, retros) and communicate progress/risks early.
  • You stay current on technical trends to suggest innovative tools and approaches to interesting problems.
  • Monitor system health using metrics, logs, and alerts, and participate in 24/7 on-call rotations to help detect, respond to, and resolve incidents.
Basic Qualifications
  • 0 to 1+ years of experience in Site Reliability Engineering, DevOps, or Platform Engineering roles
  • Hands‑on experience operating Linux‑based systems in production environments
  • Working knowledge of networking fundamentals, such as load balancing, DNS, TLS, and ingress traffic flow
  • Experience with container orchestration (e.g., EKS, Kubernetes)
  • Experience working on cloud‑native infrastructure (e.g., AWS, GCP, Azure), including networking and compute concepts
  • Proficiency in at least one programming language (e.g., Python, Ruby, Go, etc.)
  • Experience with Infrastructure as Code (e.g., Terraform, CloudFormation)
Preferred Qualifications
  • Experience with AWS cloud networking concepts such as VPCs, subnets, routing, security groups, and load balancers
  • Experience operating or contributing to production Kubernetes platforms (e.g., EKS), including cluster upgrades, networking, or ingress configuration
  • Experience with monitoring, observability, and logging platforms (e.g., DataDog, New Relic, SumoLogic, Splunk, Prometheus, Grafana)
  • Familiarity with service meshes, ingress controllers, or API gateways (e.g., Envoy, Istio, NGINX)
Salary Range

$98,000 to $148,500

Where we work

PagerDuty operates a hybrid work model with offices in 8 major cities: Atlanta, Lisbon, London, San Francisco, Santiago, Sydney, Tokyo, and Toronto. While we offer flexibility within our established locations, we cannot employ candidates residing in:

Location restrictions

Australia: Northern Territory, Queensland, South Australia, Tasmania, Western Australia; Canada: Alberta, Manitoba, Newfoundland, Northwest Territories, Nunavut, PEI, Quebec, Saskatchewan, Yukon; United States: Alaska, Hawaii, Iowa, Louisiana, Mississippi, Nebraska, New Mexico, Oklahoma, Rhode Island, South Dakota, West Virginia, Wyoming

Candidates must reside in an eligible location, which vary by role.

How We Work

Our values guide how we support customers, collaborate with colleagues, develop products, and foster a culture of belonging. They define not just our actions, but what it means to be Dutonian.

People Leaders at PagerDuty are responsible for creating high performance environments that drive accountability. PagerDuty has four key dimensions that define our Leadership Impact: Lead Self, Lead the Team, Lead the Business, and Lead the Future. Each dimension has three associated competencies to give leaders a shared language for guiding their development, career, promotion, and succession planning discussions. Our Manager Expectations serve as a practical guide for managers to understand their responsibilities, prioritize their efforts, and drive engagement and performance.

What We Offer

As a global organization, our total rewards approach is competitive with industry standards and aligned with local laws and regulations. Learn more, including country‑specific offerings, on our benefits site.

Your Package May Include
  • Competitive salary
  • Comprehensive benefits package
  • Flexible work arrangements
  • Company equity*
  • ESPP (Employee Stock Purchase Program)*
  • Retirement or pension plan*
  • Generous paid vacation time
  • Paid holidays and sick leave
  • Dutonian Wellness Days & HibernationDuty companywide paid days off in addition to PTO
  • Paid parental leave: 22 weeks for pregnant parent, 12 weeks for non‑pregnant parent (some countries have longer leave standards and we comply with local laws)*
  • Paid volunteer time off: 20 hours per year
  • Company‑wide hack weeks
  • Mental wellness programs
  • Eligibility may vary by role, region, and tenure
About PagerDuty

PagerDuty, Inc. (NYSE:PD) is a global leader in digital operations management. The PagerDuty Operations Cloud is an AI‑powered platform that empowers business resilience and drives operational efficiency for enterprises. With a generative AI assistant at its core, PagerDuty empowers teams to detect and resolve issues in real time, orchestrate complex workflows, and drive continuous improvement across their digital operations. Trusted by nearly half of both the Fortune 500 and the Forbes AI 50, as well as approximately two‑thirds of the Fortune 100, PagerDuty is essential for delivering always‑on digital experiences to modern businesses.

PagerDuty is Great Place to Work‑certified™, a Fortune Best Workplace for Millennials, a Fortune Best Medium Workplace, a Fortune Best Workplace in Technology, and a top rated product on TrustRadius and G2. Go behind‑the‑scenes on our careers site and @pagerduty on Instagram.

Additional Information

PagerDuty is an equal opportunity employer. PagerDuty does not discriminate on the basis of race, religion, color, national origin, gender, sexual orientation, age, marital status, parental status, veteran status, or disability status. Your privacy is important to us. By submitting an application, you confirm that you have read and understand PagerDuty's Privacy Policy. PagerDuty is committed to providing reasonable accommodations for qualified individuals with disabilities in our job application process. Should you require accommodation, please email accommodation@pagerduty.com and we will work with you to meet your accessibility needs. PagerDuty uses the E-Verify employment verification program.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer I
Site Reliability Engineer I

PagerDuty, Inc. • Atlanta (GA), Northern (KY)

Hybrid
USD 98,000 - 149,000
Competitive salary
Comprehensive benefits package
Company equity
+1
Site Reliability Engineer II
Site Reliability Engineer II

PagerDuty, Inc. • Atlanta (GA), Northern (KY)

Hybrid
USD 113,000 - 172,000
Competitive salary
Flexible work arrangements
Company equity
+6
Site Reliability Engineer I
Site Reliability Engineer I

Pager • Atlanta (GA), Northern (KY)

Hybrid
USD 98,000 - 149,000
Company equity
Employee Stock Purchase Program
Generous paid vacation
+2
Site Reliability Engineer I
Site Reliability Engineer I

JobCubby • Atlanta (GA), Northern (KY)

Hybrid
USD 98,000 - 149,000
Competitive salary
Benefits package
Flexible work
+10
Site Reliability Engineer II
Site Reliability Engineer II

PagerDuty • Atlanta (GA)

Hybrid
USD 113,000 - 172,000
Competitive salary
Comprehensive benefits package
Flexible work arrangements
+2
Site Reliability Engineer II
Site Reliability Engineer II

JobCubby • Atlanta (GA), Northern (KY)

Hybrid
USD 113,000 - 172,000
Competitive salary
Comprehensive benefits package
Flexible work arrangements
+6
Corporate Strategy Director
Corporate Strategy Director

PagerDuty • San Francisco (CA)

Hybrid
USD 147,000 - 246,000
Competitive salary
Comprehensive benefits package
Hybrid work model
+2
Senior AI/ML Engineer
Senior AI/ML Engineer

Triwill Group • Lisbon (ME)

Hybrid
USD 180,000 - 260,000
Equity
ESPP
Retirement plan
+7
Senior Developer Advocate
Senior Developer Advocate

Pager • Atlanta (GA)

Hybrid
USD 131,000 - 220,000
Company equity
ESPP
Retirement plan
+6
Principal Solutions Consultant
Principal Solutions Consultant

Triwill Group • United States

Hybrid
USD 180,000 - 240,000
Competitive salary
Benefits package
Flexible work arrangements
+10