Site Reliability Engineer

ASCENDING

United States

Remote

USD 150,000 - 210,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

ASCENDING is seeking a Staff Site Reliability Engineer / Cloud SME to lead the rearchitecture of a large monolithic system into a resilient, cloud-native platform across AWS and Azure.

You will own architecture transformations, drive Kubernetes-based operations (AKS/EKS), and mentor a team while advancing DevOps and security practices in a global remote environment in the continental US. The role demands deep multi-cloud expertise and a passion for scalable, secure systems.

Qualifications

  • Cloud platforms: 7+ years with AWS and Azure.
  • Cloud-native transformation: rearchitect large monoliths to cloud-native.
  • Kubernetes (AKS/EKS) expertise required.
  • Networking: design and resolve complex cloud networking issues.
  • IaC: Terraform for deployment and management.
  • Security: strong security practices for containers and Kubernetes clusters.
  • Education: Bachelor’s or Master’s in CS or related field.
  • Bonus: knowledge of load balancing algorithms.

Responsibilities

  • Lead the technical rearchitecting of a monolithic system into a microservices-based, cloud-native application.
  • Collaborate with Engineering, Architecture, and Product to define the new system using domain-driven design (DDD).
  • Conduct technology evaluations and recommend tools, frameworks, and cloud services to enhance infrastructure.
  • Develop monitoring, logging, and alerting mechanisms for reliability and availability.
  • Mentor junior team members and promote DevOps practices across the lifecycle.

Skills

AWS
Azure
Kubernetes
Cloud Networking
Security best practices
Load balancing (bonus)

Education

Bachelor's or Master's degree in Computer Science or Software Engineering

Tools

Terraform

Job description

RequiredU.S.Citizenship/Noclearanceneeded/100%remotewithintheUS
StaffSiteReliabilityEngineer/CloudSME
Location:100%remoteinthecontinentalUS
Type:Long-termcontract(3+years)

Role Summary

As the Staff SRE/Cloud SME, you will be a critical technical leader driving the rearchitecting of our existing monolithic system into a resilient, cloud-native architecture. This role requires deep expertise across multiple cloud platforms (Azure and AWS) and container orchestration (Kubernetes) to ensure the next-generation platform meets the highest standards of scalability, reliability, and security.

Key Responsibilities
Architecture & Transformation Leadership
  • Lead the technical rearchitecting efforts, transforming a large-scale monolithic system into a modern microservices-based, cloud-native application.
  • Collaborate with cross-functional teams (Engineering, Architecture, Product) to define and implement the new system architecture using domain-driven design (DDD) principles.
  • Conduct technology evaluations and provide recommendations for new tools, frameworks, and cloud services to enhance our infrastructure.
Reliability Engineering & Cloud Operations
  • Utilize Kubernetes (K8S) for container orchestration and management, ensuring extreme scalability, reliability, and high availability of the system.
  • Implement robust, highly resilient, and highly available components for the system.
  • Develop and implement comprehensive monitoring, logging, and alerting mechanisms to ensure optimal system performance and availability.
  • Drive the adoption of DevOps principles and practices throughout the software development lifecycle, ensuring seamless integration and continuous deployment processes.
Technical Expertise & Mentorship
  • Stay up-to-date with emerging technologies, frameworks, and industry trends related to systems and cloud computing.
  • Mentor and provide technical guidance to junior team members, fostering a culture of continuous learning and professional growth.
Required Qualifications
  • Cloud Platforms: 7+ years of experience with cloud computing platforms. Strong multi-cloud expertise required with AWS and Azure.
  • Cloud-Native Transformation: 7+ years of experience in rearchitecting large-scale monolithic applications to cloud-native architectures.
  • Container Orchestration: Strong expertise in Kubernetes (K8S) is required, including hands-on experience with both AKS (Azure Kubernetes Service) and EKS (Elastic Kubernetes Service).
  • Networking: Strong experience with Cloud Networking, with the ability to design and resolve complex cloud networking architecture problems.
  • IaC: Expert knowledge of Terraform for infrastructure-as-code deployment and management.
  • Security: Must possess strong knowledge of security best practices for containers and Kubernetes clusters.
  • Education: Bachelor's or Master's degree in Computer Science, Software Engineering, or a related field.
  • Bonus Knowledge: Knowledge of load balancing algorithms.

Thanksforapplying!

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Staff Site Reliability Engineer
Senior Staff Site Reliability Engineer

Archer • San Jose (CA)

On-site
USD 160,000 - 210,000
Remote Staff SRE: Cloud-Native Transformation Lead
Remote Staff SRE: Cloud-Native Transformation Lead

ascendingdc • Fairfax (VA)

Remote
USD 120,000 - 160,000
Staff SRE & Cloud Architect (Remote, US)
Staff SRE & Cloud Architect (Remote, US)

ASCENDING • United States

Remote
USD 150,000 - 210,000
Senior SRE & Cloud Architect (AWS/Azure) – Remote
Senior SRE & Cloud Architect (AWS/Azure) – Remote

ASCENDING • Austin (TX)

On-site
USD 160,000 - 210,000
Senior/Staff Cloud Reliability Engineer
Senior/Staff Cloud Reliability Engineer

Cerebras • Mountain View (CA)

On-site
USD 190,000 - 240,000
Senior Software Engineer- Site Reliability Engineering (SRE)
Senior Software Engineer- Site Reliability Engineering (SRE)

Noctua Technology • Virginia (MN), California (MO), Washington

Remote
USD 149,000 - 202,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mike Albert Fleet Solutions • Cincinnati (OH)

Hybrid
USD 100,000 - 135,000
Site Reliability Engineering Manager
Site Reliability Engineering Manager

Calance • United States

Hybrid
USD 150,000 - 200,000
Site Reliability Engineer
Site Reliability Engineer

Evlo AI • Minneapolis (MN)

On-site
USD 120,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Harrison Clarke • New York (NY)

On-site
USD 120,000 - 160,000