Site Reliability Engineer (GKE) - Draper, UT

Wasatch Property Management, Inc.

United States

On-site

USD 112,500 - 137,500

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Comprehensive medical, dental, and vision coverage
401(k) with company match
3-5 weeks PTO annually
Paid company holidays

Job summary

Wasatch Property Management, Inc. is looking for a hands-on Site Reliability Engineer to enhance their infrastructure operations. The successful candidate will bridge software development and operations, ensuring optimal performance and reliability of the cloud infrastructure.

The role includes designing resilient systems, optimizing CI/CD pipelines, and conducting incident responses. A degree in Computer Science or related fields and 3+ years of cloud experience are required, along with a competitive compensation package and benefits.

Qualifications

  • 3+ years of experience in an SRE, DevOps, or Infrastructure Engineering role.
  • Hands-on experience with cloud providers (AWS or GCP).
  • Strong programming proficiency in automation languages.

Responsibilities

  • Design and maintain scalable multi-tenant cloud infrastructure.
  • Own the uptime and performance of cloud platforms.
  • Develop robust monitoring and alerting systems.

Skills

Site Reliability Engineering
Cloud Architecture
Automation
Incident Response
Programming (Python, Go, TypeScript, Bash)
Containerization (Docker, Kubernetes)
Networking Protocols

Education

Bachelor's or Master's degree in Computer Science, Information Systems, or related fields

Tools

AWS
GCP
GitHub Actions
Argo Workflows
Prometheus
Grafana

Job description

Role Overview

We are seeking a hands‑on Site Reliability Engineer to join our team. You will bridge the gap between software development and infrastructure operations, treating operational challenges as engineering problems. By leveraging automation, designing resilient distributed systems, and championing observability, you will ensure that our customers can secure and manage their access control systems without friction or failure.

Key Responsibilities
  • Infrastructure & Automation: Design, build, and maintain scalable, secure multi‑tenant cloud infrastructure using Infrastructure as Code (IaC) principles.
  • Uptime & Reliability: Own the availability, latency, performance, and capacity planning of the pdk.io platform and its supporting backend microservices.
  • Observability: Develop and manage robust monitoring, logging, and alerting systems to gain deep visibility into cloud infrastructure, API health, and IoT endpoint performance.
  • Incident Response: Participate in a collaborative on‑call rotation, lead rapid incident response mitigation and drive rigorous, blameless post‑mortems to ensure long‑term system resilience.
  • CI/CD Pipeline Management: Optimize and secure automated deployment pipelines to enable developers to ship code to production safely and efficiently.
  • Cross‑Functional Collaboration: Partner closely with backend developers and hardware engineering teams to define Service Level Indicators (SLIs), Service Level Objectives (SLOs), and manage error budgets. Contribute to technical documentation and knowledge sharing.
  • Tooling: CI automation with GitHub Actions and Argo Workflows; IaC with OpenTofu & Terragrunt; GitOps continuous delivery using ArgoCD; Observability stack powered by Prometheus and Grafana.
Required Qualifications
  • Bachelor’s or Master’s degree in Computer Science, Information Systems, or related fields or equivalent experience.
  • 3+ years of experience in an SRE, DevOps, or Infrastructure Engineering role supporting production cloud environments.
  • Cloud Architecture: Deep hands‑on experience with major cloud providers (AWS or GCP) and a strong command of containerization and orchestration technologies (Docker and Kubernetes).
  • Coding & Scripting: Strong programming proficiency in languages such as Python, Go, TypeScript, or Bash for automation, internal tooling, and system integrations.
  • Systems & Networking: Solid fundamentals in Linux/Unix administration, networking protocols (TCP/IP, DNS, HTTP/S, load balancing), and cloud security best practices.
  • Mindset: A passionate problem‑solver who prioritizes automation over manual operations and thrives in high‑ownership environments.
  • Must pass drug and criminal background check.
  • Work well in an onsite team environment.
Preferred Qualifications
  • Experience with multi‑region deployments, failover strategies, and data consistency.
  • Messaging Systems: Experience managing high‑throughput message queues or data streaming platforms, specifically RabbitMQ or Apache Kafka.
  • Experience operating production systems at scale.
  • Data Infrastructure: Familiarity with modern data stack environments, such as Snowflake or relational databases in a self‑hosted environment.
  • Certifications: Relevant industry certifications such as Certified Kubernetes Administrator (CKA) or AWS Certified DevOps Engineer Professional.
  • Familiarity with regulatory requirements (SOC2, GDPR, etc).
  • Nice to Have: Experience in physical security or access control systems.
  • Familiarity with GCP ecosystem and tooling.
  • Experience working in a scaling startup environment.
Compensation & Benefits
  • Competitive salary starting at $125,000 depending on experience.
  • Comprehensive medical, dental, and vision coverage.
  • 401(k) with company match.
  • 3‑5 weeks PTO annually based on tenure.
  • Paid company holidays.
Work Location

This position is an in‑office, non‑remote role, working from ProdataKey Headquarters located in Draper, Utah.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

ProdataKey • Draper (UT)

On-site
USD 75,000 - 125,000
Comprehensive medical coverage
Dental and vision coverage
401(k) with company match
+2
Site Reliability Engineer (GKE) - Draper, UT
Site Reliability Engineer (GKE) - Draper, UT

Wasatch Property Management, Inc. • Draper (UT)

On-site
USD 125,000 - 145,000
Comprehensive medical, dental, and vision coverage
401(k) with company match
3-5 weeks PTO annually based on tenure
+1
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • North Carolina

On-site
USD 165,000 - 215,000
Pre‑IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Site Reliability Engineer - Build Resilient Cloud Infra
Site Reliability Engineer - Build Resilient Cloud Infra

ProdataKey • Draper (UT)

On-site
USD 75,000 - 125,000
Comprehensive medical coverage
Dental and vision coverage
401(k) with company match
+2
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New York (NY)

Hybrid
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+1
Senior DevOps Engineer/Site Reliability Engineer-East Coast
Senior DevOps Engineer/Site Reliability Engineer-East Coast

Stellar Cyber • New Jersey

On-site
USD 165,000 - 215,000
Pre-IPO Stock Options
Medical, Dental & Vision care
401(k)
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Drata • San Francisco (CA)

Hybrid
USD 166,000 - 226,000
Stock equity
Up to 100% employer-paid medical coverage
401(k) plan
+2
Senior Site Reliability Engineer
Senior Site Reliability Engineer

O.C. Tanner • Salt Lake City (UT)

On-site
USD 130,000 - 180,000
Senior Site Reliability Engineer- Sunnyvale, CA, the US
Senior Site Reliability Engineer- Sunnyvale, CA, the US

Kody • Sunnyvale (CA)

On-site
USD 120,000 - 160,000
Competitive package
Collaborative, inclusive environment
Site Reliability Engineer
Site Reliability Engineer

Harrison Clarke • New York (NY)

On-site
USD 120,000 - 160,000