Senior SRE - Azure

Compunnel, Inc.

Alpharetta (GA)

On-site

USD 100,000 - 140,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

A leading technology firm is seeking a Senior Site Reliability Engineer (Azure) to build, maintain, and scale cloud-native infrastructure. This role will work with development and operations teams to ensure reliable and efficient systems. Key responsibilities include designing Azure environments, managing Kubernetes clusters, and enhancing CI/CD pipelines using Terraform and Databricks. The ideal candidate will have 4+ years in Site Reliability Engineering or cloud infrastructure, strong Azure skills, and a collaborative mindset.

Qualifications

  • Minimum 4 years of experience in Site Reliability Engineering, DevOps, or cloud infrastructure roles.
  • Strong hands-on experience with Azure cloud services.
  • Proficiency with Java and Infrastructure-as-Code tools including Terraform and Terragrunt.
  • Strong experience with Kubernetes (preferably AKS) and container orchestration.
  • Experience working with Databricks in production environments.
  • Proficiency with CI/CD tooling, especially GitHub Workflows/Actions and ArgoCD.
  • Strong understanding of observability tooling, including Grafana (Prometheus, Loki, Tempo preferred).
  • Ability to collaborate in cross-functional environments and communicate effectively.

Responsibilities

  • Design, implement, and maintain Azure cloud infrastructure using best practices.
  • Manage and optimize Kubernetes clusters and containerized workloads.
  • Build and maintain Infrastructure-as-Code solutions.
  • Develop, maintain, and enhance CI/CD pipelines.
  • Support Databricks environments and associated integrations.
  • Implement and improve observability using various tools.
  • Automate operational tasks to improve efficiency.
  • Participate in on-call rotations, incident response, and root-cause analysis.
  • Collaborate with developers to improve application performance.
  • Identify opportunities for cost optimization and infrastructure security enhancements.

Skills

Site Reliability Engineering
DevOps
Azure cloud services
Kubernetes
Terraform
CI/CD
Observability tools

Education

Master’s degree in Computer Science or related field

Tools

Terraform
Terragrunt
GitHub Workflows/Actions
ArgoCD
Grafana
Prometheus
Loki
Tempo
Databricks

Job description

The Senior Site Reliability Engineer (Azure) will help build, maintain, and scale cloud-native infrastructure in a fast-paced, collaborative environment.

This role works closely with development and operations teams to ensure systems are reliable, efficient, automated, and secure.

Responsibilities include designing Azure cloud environments, managing Kubernetes clusters, implementing Infrastructure-as-Code through Terraform/Terragrunt, improving CI/CD pipelines, enhancing system observability, and providing support for on-call and incident response activities.

Key Responsibilities
  • Design, implement, and maintain Azure cloud infrastructure using best practices for scalability and reliability.
  • Manage and optimize Kubernetes clusters (preferably AKS) and containerized workloads.
  • Build and maintain Infrastructure-as-Code solutions using Terraform and Terragrunt.
  • Develop, maintain, and enhance CI/CD pipelines using GitHub Workflows/Actions and ArgoCD.
  • Support Databricks environments and associated cloud integrations.
  • Implement and improve observability using tools such as Grafana, Prometheus, Loki, and Tempo.
  • Automate operational tasks to improve efficiency, reduce manual work, and enhance reliability.
  • Participate in on-call rotations, incident response, root-cause analysis, and remediation activities.
  • Collaborate with developers to improve application performance, reliability, and adherence to SRE practices such as SLIs and SLOs.
  • Identify opportunities for cost optimization, performance improvements, and infrastructure security enhancements.
Required Qualifications
  • Minimum 4 years of experience in Site Reliability Engineering, DevOps, or cloud infrastructure roles.
  • Strong hands‑on experience with Azure cloud services.
  • Proficiency with Java and Infrastructure-as-Code tools including Terraform and Terragrunt.
  • Strong experience with Kubernetes (preferably AKS) and container orchestration.
  • Experience working with Databricks in production environments.
  • Proficiency with CI/CD tooling, especially GitHub Workflows/Actions and ArgoCD.
  • Strong understanding of observability tooling, including Grafana (Prometheus, Loki, Tempo preferred).
  • Ability to collaborate in cross‑functional environments and communicate effectively.
Preferred Qualifications
  • Master’s degree in Computer Science or a related field.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mike Albert Fleet Solutions • Cincinnati (OH)

Hybrid
USD 100,000 - 135,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Bank of America • Chandler (AZ)

On-site
USD 140,000 - 190,000
Sr. Site Reliability Engineer- THADC5716474
Sr. Site Reliability Engineer- THADC5716474

Compunnel Inc. • Alpharetta (GA)

On-site
USD 90,000 - 120,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mikealbert • Cincinnati (OH)

Hybrid
USD 100,000 - 130,000
Site Reliability Engineer
Site Reliability Engineer

SPECTRAFORCE • Alpharetta (GA)

On-site
USD 90,000 - 120,000
Senior SRE Engineer Azure Healthcare Observability Healthcare SaaS
Senior SRE Engineer Azure Healthcare Observability Healthcare SaaS

AppRecode, Inc. • Town of Middletown (NY)

On-site
USD 120,000 - 150,000
Site Reliability Engineer
Site Reliability Engineer

NextGen | GTA: A Kelly Telecom Company • Mount Laurel Township (NJ)

On-site
USD 110,000 - 170,000
Senior SRE Engineer
Senior SRE Engineer

Compunnel, Inc. • Alpharetta (GA)

On-site
USD 140,000 - 190,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Koitecc Solutions • Plano (TX), Northern (KY)

Hybrid
USD 153,000 - 192,000
Discretionary incentive
Benefits package
Azure SRE Architect: Scale, Automate & Observability
Azure SRE Architect: Scale, Automate & Observability

Compunnel, Inc. • Alpharetta (GA)

On-site
USD 100,000 - 140,000