Senior SRE Architect: Cloud Reliability

Socket.dev

Denver (CO)

Hybrid

USD 110,000 - 146,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Medical, dental and vision coverage
401(k) retirement savings options
Paid holidays, vacation time and sick,
Travel privileges on Frontier Airlines
Buddy passes
Travel-related discounts
Hybrid schedule for HQ roles in Denver

Job summary

Frontier Airlines, a Denver-based carrier, is seeking a Lead Site Reliability Engineer to drive the reliability, scalability, and performance of mission-critical platforms. You will partner with engineering, platform, security and operations teams to build resilient systems and advocate reliability-driven engineering practices across the organization.

You should bring deep expertise in AWS, Kubernetes, observability, and DevSecOps, with a track record of leading large-scale cloud migrations,

Qualifications

  • Bachelor’s degree in computer science, engineering, information technology, or related field.
  • 10+ years of experience in software engineering, cloud architecture, infrastructure engineering, or enterprise architecture.
  • 5+ years of hands-on AWS architecture and cloud transformation experience.
  • Proven success leading large-scale cloud migrations and modernization initiatives.
  • Experience designing and supporting highly available, mission-critical platforms.
  • Deep understanding of Kubernetes, container orchestration, microservices, and distributed systems.
  • Extensive experience with DevSecOps, CI/CD pipelines, Infrastructure as Code, and automation frameworks.
  • Strong knowledge of Site Reliability Engineering (SRE), operational excellence, and platform reliability practices.
  • Experience implementing cloud governance, FinOps, and cost optimization programs.

Responsibilities

  • Lead design of highly available, resilient cloud platforms using AWS.
  • Define reliability strategies with SLOs/SLIs and error budgets for critical services.
  • Architect Kubernetes-based platforms supporting containerized apps.
  • Establish enterprise standards for reliability, availability, and disaster recovery.
  • Partner with engineering, security, and operations to improve service reliability.
  • Drive DevSecOps adoption, CI/CD automation, and Infrastructure as Code.
  • Define observability standards: monitoring, logging, tracing, and analytics.
  • Lead incident response, RCAs, and continuous improvement initiatives.
  • Eliminate single points of failure through proactive reliability engineering.
  • Design automated remediation and self-healing capabilities.
  • Lead capacity planning, performance engineering, and scalability assessments.
  • Collaborate on FinOps to optimize cloud resource use.
  • Evaluate AIOps and generative AI to improve reliability.
  • Mentor SREs and engineers in reliability practices.

Skills

Leadership
Operational strategy
Executive communication
Cross-functional collaboration

Education

Bachelor's degree in CS/Engineering/IT

Tools

AWS
Kubernetes (EKS)
Terraform
CloudFormation
CI/CD pipelines
Monitoring tools

Job description

Frontier Airlines, a Denver-based carrier, is seeking a Lead Site Reliability Engineer to drive the reliability, scalability, and performance of mission-critical platforms. You will partner with engineering, platform, security and operations teams to build resilient systems and advocate reliability-driven engineering practices across the organization.

You should bring deep expertise in AWS, Kubernetes, observability, and DevSecOps, with a track record of leading large-scale cloud migrations,

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE - Cloud Reliability & Automation
Senior SRE - Cloud Reliability & Automation

Frontier-Airlines • Denver (CO)

Hybrid
USD 110,000 - 146,000
Medical/Dental/Vision
401(k) retirement plan
Travel privileges
+3
Senior SRE - Cloud & Observability
Senior SRE - Cloud & Observability

Ridgeline • Reno (NV)

Hybrid
USD 153,000 - 210,000
Unlimited vacation
Education reimbursement
Wellness reimbursement
+1
Senior SRE & Cloud Reliability Architect
Senior SRE & Cloud Reliability Architect

United States Digital Space LLC • United States

Remote
USD 180,000 - 250,000
Senior SRE - Remote Cloud Platform Reliability & Automation
Senior SRE - Remote Cloud Platform Reliability & Automation

Piper Companies • United States

Remote
USD 130,000 - 180,000
Medical, dental, vision coverage
401(k)
Paid time off
+1
Senior SRE: Architect Scalable, Reliable Cloud Infra
Senior SRE: Architect Scalable, Reliable Cloud Infra

The ReWork Group • New York (NY)

On-site
USD 120,000 - 160,000
SRE Architecture Lead: Reliability & Cloud Platform
SRE Architecture Lead: Reliability & Cloud Platform

MACHINE LEARNING TECHNOLOGIES LLC • Atlanta (GA)

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer - Cloud Platform & Automation
Senior Site Reliability Engineer - Cloud Platform & Automation

Ridgeline, Inc. • San Ramon (CA)

Hybrid
USD 153,000 - 210,000
Unlimited vacation
Educational reimbursement
Comprehensive insurance plans
Lead SRE - Remote/Hybrid, Enterprise-Scale Reliability
Lead SRE - Remote/Hybrid, Enterprise-Scale Reliability

Empower Retirement • Greenwood Village (CO)

Hybrid
USD 114,000 - 166,000
Medical insurance
401(k) with company match
Tuition reimbursement
+3
SRE Manager: Reliability Leader for Scalable Cloud
SRE Manager: Reliability Leader for Scalable Cloud

Litera • Denver (CO)

Hybrid
USD 120,000 - 160,000
Senior Cloud SRE - Platform Reliability & Automation
Senior Cloud SRE - Platform Reliability & Automation

Carrier Transicold Polska Sp. z o.o. • Northern (KY)

Hybrid
USD 96,000 - 192,000
Health Care Benefits
Retirement Benefits
Paid vacation days and holidays