Azure SRE Architect: Scale, Automate & Observability

Compunnel, Inc.

Alpharetta (GA)

On-site

USD 100,000 - 140,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A leading technology firm is seeking a Senior Site Reliability Engineer (Azure) to build, maintain, and scale cloud-native infrastructure. This role will work with development and operations teams to ensure reliable and efficient systems. Key responsibilities include designing Azure environments, managing Kubernetes clusters, and enhancing CI/CD pipelines using Terraform and Databricks. The ideal candidate will have 4+ years in Site Reliability Engineering or cloud infrastructure, strong Azure skills, and a collaborative mindset.

Qualifications

  • Minimum 4 years of experience in Site Reliability Engineering, DevOps, or cloud infrastructure roles.
  • Strong hands-on experience with Azure cloud services.
  • Proficiency with Java and Infrastructure-as-Code tools including Terraform and Terragrunt.
  • Strong experience with Kubernetes (preferably AKS) and container orchestration.
  • Experience working with Databricks in production environments.
  • Proficiency with CI/CD tooling, especially GitHub Workflows/Actions and ArgoCD.
  • Strong understanding of observability tooling, including Grafana (Prometheus, Loki, Tempo preferred).
  • Ability to collaborate in cross-functional environments and communicate effectively.

Responsibilities

  • Design, implement, and maintain Azure cloud infrastructure using best practices.
  • Manage and optimize Kubernetes clusters and containerized workloads.
  • Build and maintain Infrastructure-as-Code solutions.
  • Develop, maintain, and enhance CI/CD pipelines.
  • Support Databricks environments and associated integrations.
  • Implement and improve observability using various tools.
  • Automate operational tasks to improve efficiency.
  • Participate in on-call rotations, incident response, and root-cause analysis.
  • Collaborate with developers to improve application performance.
  • Identify opportunities for cost optimization and infrastructure security enhancements.

Skills

Site Reliability Engineering
DevOps
Azure cloud services
Kubernetes
Terraform
CI/CD
Observability tools

Education

Master’s degree in Computer Science or related field

Tools

Terraform
Terragrunt
GitHub Workflows/Actions
ArgoCD
Grafana
Prometheus
Loki
Tempo
Databricks

Job description

A leading technology firm is seeking a Senior Site Reliability Engineer (Azure) to build, maintain, and scale cloud-native infrastructure. This role will work with development and operations teams to ensure reliable and efficient systems. Key responsibilities include designing Azure environments, managing Kubernetes clusters, and enhancing CI/CD pipelines using Terraform and Databricks. The ideal candidate will have 4+ years in Site Reliability Engineering or cloud infrastructure, strong Azure skills, and a collaborative mindset.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Azure SRE: Cloud Automation, CI/CD & Observability
Azure SRE: Cloud Automation, CI/CD & Observability

SPECTRAFORCE • Alpharetta (GA)

On-site
USD 90,000 - 120,000
Azure SRE Leader: Observability, Automation & Resilience
Azure SRE Leader: Observability, Automation & Resilience

TechDigital Group • Dallas (TX)

On-site
USD 100,000 - 130,000
Azure SRE: Build Resilient, High-Performance Cloud Systems
Azure SRE: Build Resilient, High-Performance Cloud Systems

EBSCO Health • Birmingham (AL)

On-site
USD 100,000 - 130,000
Azure Cloud SRE Lead — Migration, CI/CD & Observability
Azure Cloud SRE Lead — Migration, CI/CD & Observability

Systems Technology Group, Inc. (STG) • Dearborn (MI)

On-site
USD 100,000 - 130,000
Azure SRE Engineer - Cloud Infra, GitOps & Observability
Azure SRE Engineer - Cloud Infra, GitOps & Observability

Motion Recruitment • Alpharetta (GA)

Hybrid
USD 100,000 - 125,000
Azure SRE: Greenfield Distributed Systems Architect
Azure SRE: Greenfield Distributed Systems Architect

Storm2 • United States

Remote
USD 120,000 - 150,000
Azure SRE: Reliability, Observability & Incident Leadership
Azure SRE: Reliability, Observability & Incident Leadership

Veriipro • Deerfield (IL)

On-site
Senior Azure SRE — Remote, Automation‑First Reliability
Senior Azure SRE — Remote, Automation‑First Reliability

Concord Technologies • United States

Remote
USD 120,000 - 160,000
Azure Cloud Engineer - SRE, CI/CD & Kubernetes
Azure Cloud Engineer - SRE, CI/CD & Kubernetes

GPRS • Kentucky

On-site
USD 100,000 - 130,000
Senior Azure SRE — Remote (EST)
Senior Azure SRE — Remote (EST)

Korn Ferry • Philadelphia

Remote
USD 100,000 - 125,000