SRE-Azure - Global Industrial

Motion

Birmingham, Northern (AL, KY)

Hybrid

USD 110,000 - 170,000

Full time

2 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Healthcare coverage
401(k)
Tuition reimbursement
Vacation
Sick leave
Holiday pay

Job summary

Motion is seeking an Azure Site Reliability Engineer (SRE) to enhance reliability, scalability, and performance of enterprise apps on Microsoft Azure. The role combines software engineering, cloud infrastructure, automation, and DevOps practices to reduce toil and improve resilience.

You will work with AKS, Terraform, ARM Templates, and monitoring tools, collaborating with development, cloud, architecture, and security teams to drive reliability initiatives and production readiness.

Qualifications

  • Bachelor's degree and 5–7 years in tech/software engineering.
  • Experience with large-scale, highly available cloud platforms.
  • Strong knowledge of SRE, incident management, RCAs, and observability.

Responsibilities

  • Monitor and optimize Azure-based platforms for reliability and performance.
  • Define SLIs, SLOs, SLAs and error budgets; manage service reliability.
  • Automate deployments and runbooks; support incident response and RCA.
  • Work with AKS, Azure Networking, App Services, and Storage.
  • Participate in on-call rotations and improve deployment pipelines.

Skills

SRE principles
Incident management
Root cause analysis
Communication

Education

Bachelor's degree

Tools

Terraform
ARM Templates
AKS
Kubernetes
GitHub Actions
Azure DevOps

Job description

## SRE-Azure - Global IndustrialApply: Hybrid: Birmingham, AL, USA: Atlanta, GA, USA: Full time: Posted Yesterday: R26\\_0000031014**The Azure Site Reliability Engineer (SRE)** is responsible for improving the reliability, availability, scalability, performance, and operational excellence of enterprise applications and platforms hosted within Microsoft Azure. This role combines software engineering, cloud infrastructure, automation, and DevOps practices to build and support resilient, cloud-native solutions while reducing operational toil through automation.The Azure SRE leverages Azure platform services, Kubernetes, Infrastructure as Code (IaC), and observability tools to ensure mission-critical systems remain highly available, secure, and performant. This role partners closely with application development, cloud engineering, architecture, and cybersecurity teams to drive continuous improvement, accelerate cloud adoption, and maintain service reliability through proactive monitoring, incident management, and operational engineering.**JOB DUTIES*** Monitor, analyze, and optimize system performance, availability, and reliability across Azure-hosted platforms and applications.* Define and manage Service Level Indicators (SLIs), Service Level Objectives (SLOs), Service Level Agreements (SLAs), and Error Budgets.* Drive continuous service improvement through operational metrics, trend analysis, reliability engineering practices, and platform modernization efforts.* Partner with development teams to improve service reliability through testing, release validation, deployment automation, and production readiness reviews.* Design, build, and support Azure infrastructure and platform services, including Azure Kubernetes Service (AKS), Azure Networking, App Services, and Storage* Develop and maintain Infrastructure as Code (IaC) solutions utilizing Terraform and Azure-native deployment technologies.* Lead and support incident response, root cause analysis (RCA), post-incident reviews, and service restoration efforts.* Automate operational processes, platform provisioning, deployments, and remediation activities to reduce manual effort (TOIL) and improve reliability.* Identify, investigate, and mitigate performance, security, networking, and availability issues, including traffic anomalies and service disruptions.* Participate in on-call rotations and maintain operational documentation, runbooks, and knowledge articles as needed.**EDUCATION & EXPERIENCE*** Typically requires a bachelor's degree and five (5) to seven (7) years of experience in a technology and/or software engineering role or an equivalent combination.* Experience supporting large-scale, highly available, distributed applications and cloud platforms.**KNOWLEDGE, SKILLS, ABILITIES*** Strong understanding of Site Reliability Engineering principles, including SLIs, SLOs, Error Budgets, Incident Management, Root Cause Analysis, and Operational Excellence.* Hands-on experience with Microsoft Azure services, including infrastructure, networking, security, platform services, and cloud-native technologies.* Experience administering and supporting Azure Kubernetes Service (AKS), Kubernetes clusters, containers, and scalable distributed systems.* Proficiency with Infrastructure as Code (Terraform, ARM Templates) and Git-based deployment practices.* Experience with Azure DevOps, GitHub Actions, CI/CD pipelines, release automation, and DevOps methodologies.* Strong troubleshooting skills across cloud infrastructure, operating systems, databases, networking, and security domains.* Experience with monitoring and observability platforms, including Azure Monitor, Log Analytics, Application Insights, Grafana, Datadog, and Dynatrace.* Knowledge of microservices, APIs, distributed architectures, and cloud-native design patterns.* Working knowledge of Windows Server, Linux, networking, DNS, load balancing, firewalls, and hybrid cloud connectivity.* Experience with capacity planning, performance engineering, scalability testing, disaster recovery, and business continuity practices.* Strong analytical, problem-solving, communication, and collaboration skills. **PHYSICAL DEMANDS:****LICENSES & CERTIFICATIONS:****SUPERVISORY RESPONSIBILITY:****BUDGET RESPONSIBILITY:** No **COMPANY INFORMATION:** Motion offers an excellent benefits package which includes options for healthcare coverage, 401(k), tuition reimbursement, vacation, sick, and holiday pay. **DISCLAIMER:** This job description illustrates the general nature and level of work performed by employees within this job classification. It is not intended to contain or be interpreted as a comprehensive inventory of all duties, responsibilities and skills required. Management retains the right to add or modify duties at any time.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

SRE-Azure - Global Industrial
SRE-Azure - Global Industrial

Motion Industries (MOT) • Alabama

On-site
USD 110,000 - 160,000
Healthcare coverage
401(k) plan
Tuition reimbursement
SRE-Azure - Global Industrial
SRE-Azure - Global Industrial

Genuine Parts Company • Alabama

On-site
USD 110,000 - 160,000
Healthcare coverage
401(k)
Tuition reimbursement
+3
Senior SRE - Azure
Senior SRE - Azure

Compunnel, Inc. • Alpharetta (GA)

On-site
USD 100,000 - 140,000
Senior System Reliability Engineer
Senior System Reliability Engineer

On-Demand Group • Eagan (MN)

On-site
USD 229,233,000 - 286,541,000
Reliability Engineer (52718)
Reliability Engineer (52718)

GAP Solutions, Inc. • Atlanta (GA)

Hybrid
USD 90,000 - 140,000
Site Reliability Engineer (SRE)
Site Reliability Engineer (SRE)

Nuvento Inc • Kansas City (MO)

On-site
USD 120,000 - 160,000
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Mike Albert Fleet Solutions • Cincinnati (OH)

On-site
USD 100,000 - 135,000
Site Reliability Engineer
Site Reliability Engineer

Moultrie • Birmingham (AL)

On-site
USD 110,000 - 170,000
Azure SRE: Cloud-Native Reliability Engineer
Azure SRE: Cloud-Native Reliability Engineer

Motion • Birmingham (AL), Northern (KY)

Hybrid
USD 110,000 - 170,000
Healthcare coverage
401(k)
Tuition reimbursement
+3
SRE-Azure - Global Industrial
SRE-Azure - Global Industrial

USA MOT Motion Industries, Inc. • Alton (IL)

On-site
USD 120,000 - 170,000
Healthcare coverage
401(k)
Tuition reimbursement
+3