Azure Cloud SRE & Reliability Engineer

Graphnet Health Ltd.

United States

Remote

USD 120,000 - 180,000

Full time

44 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Graphnet Health Ltd. is seeking a Site Reliability Engineer to join the TechOps Team and own reliability of the CareCentric platform. The role involves monitoring, incident response, and collaboration with Operations, Security, Development, and Project Teams.

You will leverage Azure services, implement observability, and automate repetitive tasks. On-call rotation and continuous improvement are part of the job.

Qualifications

  • Minimum of four years working within an SRE function.
  • Experience of service monitoring and alerting.
  • Azure PaaS components including AKS, Application Insights / Log Analytics, App Services, Azure SQL, Storage.
  • Networking (NSG, VLANs etc).
  • Proven experience with Azure DevOps (ADO) and/or GitHub.
  • Experience in the support of Cloudflare.
  • Understanding of Terraform or other IAC tooling.
  • Experience of routine infrastructure upgrades and platform patching.
  • Demonstrable PowerShell administration / scripting skills.
  • Experience of Service Desk Systems.

Responsibilities

  • Provision of first-class infrastructure support to customers and systems for all Managed Service Platforms.
  • Implement observability and proactively monitor our Azure Cloud footprint, utilising dashboards, alerts and runbooks.
  • Incident response, analysis, remediation and associated post-incident root cause analysis reviews.
  • Support performance / reliability testing and capacity planning activities.
  • Perform routine system upgrades and platform patching.
  • Promote automation of repetitive tasks, reducing manual intervention.
  • Eventual participation in the 24x7 On-Call Rota.
  • Excellent, demonstrable troubleshooting and problem-solving skills.
  • Able to work well as an individual and as part of a team.
  • Take part in architectural discussions with technical / non-technical team members.
  • Organisational skills, with the ability to manage personal workloads in accordance with agreed timescales, whilst working under pressure.
  • An eye for detail and a desire to adhere to best practices.
  • Strong inter-personal and communication skills.
  • Have a desire to keep up with the latest tools and techniques.
  • Verbal, written communication and documentational skills.
  • Ability to explain technical concepts to key stakeholders of all levels.

Skills

SRE experience
Monitoring & alerting
Azure
AKS
Application Insights / Log Analytics
App Services
Azure SQL
Networking
Azure DevOps / GitHub
Terraform / IAC
PowerShell
Service Desk Systems

Tools

Azure DevOps (ADO)
GitHub
Terraform
PowerShell

Job description

Graphnet Health Ltd. is seeking a Site Reliability Engineer to join the TechOps Team and own reliability of the CareCentric platform. The role involves monitoring, incident response, and collaboration with Operations, Security, Development, and Project Teams.

You will leverage Azure services, implement observability, and automate repetitive tasks. On-call rotation and continuous improvement are part of the job.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Graphnet Health Ltd. • United States

Remote
USD 120,000 - 180,000
Azure SRE: Cloud-Native Reliability Engineer
Azure SRE: Cloud-Native Reliability Engineer

Motion • Birmingham (AL), Northern (KY)

Hybrid
USD 110,000 - 170,000
Healthcare coverage
401(k)
Tuition reimbursement
+3
Senior Azure SRE — Remote, Automation‑First Reliability
Senior Azure SRE — Remote, Automation‑First Reliability

Concord Technologies • United States

Remote
USD 120,000 - 160,000
401K plan with 6% company match
Flex-time off
Paid parental leave
+3
Azure SRE Lead: Production Reliability & Cloud Ops
Azure SRE Lead: Production Reliability & Cloud Ops

Experis Technology Group • Alpharetta (GA)

On-site
USD 72,000 - 96,000
Medical and Prescription Plans
Dental Plan
Vision Plan
+4
Azure SRE Architect: Scale, Automate & Observability
Azure SRE Architect: Scale, Automate & Observability

Compunnel, Inc. • Alpharetta (GA)

On-site
USD 100,000 - 140,000
Site Reliability Engineer
Site Reliability Engineer

Moultrie • Birmingham (AL)

On-site
USD 110,000 - 170,000
Remote Azure SRE – Observability, CI/CD & Resilience
Remote Azure SRE – Observability, CI/CD & Resilience

System Automation Corporation • United States

On-site
USD 120,000 - 140,000
Azure SRE: Cloud Reliability & Automation Engineer
Azure SRE: Cloud Reliability & Automation Engineer

Genuine Parts Company • Alabama

On-site
USD 110,000 - 160,000
Healthcare coverage
401(k)
Tuition reimbursement
+3
Azure Site Reliability Engineer | Reliability & Automation
Azure Site Reliability Engineer | Reliability & Automation

Ebsco Subscription Services España SL • Birmingham (AL)

On-site
USD 90,000 - 120,000
Azure Cloud SRE Lead — Migration, CI/CD & Observability
Azure Cloud SRE Lead — Migration, CI/CD & Observability

Systems Technology Group, Inc. (STG) • Dearborn (MI)

On-site
USD 100,000 - 130,000