Site Reliability Engineer

Graphnet Health Ltd.

Milton Keynes

Hybrid

GBP 70,000 - 95,000

Full time

3 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Graphnet Health Ltd. in Milton Keynes is seeking a Site Reliability Engineer (SRE) to join the TechOps team. You will own infrastructure reliability, monitoring, and incident response for the CareCentric platform.

You will work with Operations, Security, Development and Project teams, gain on-the-job training, and be comfortable with Azure, Windows Server, networks, and Terraform/IAC tools. The role includes observability setup, patching, capacity planning, and potential 24x7 on-call rotation,

Qualifications

  • Minimum of four years working within an SRE function.
  • Experience of service monitoring and alerting.
  • Microsoft Azure PaaS components including Storage.
  • Proven experience with Azure DevOps (ADO) and/or GitHub.
  • Experience in the support of Cloudflare.
  • An understanding of Terraform or other IAC tooling.
  • Experience of routine infrastructure upgrades and platform patching.
  • Demonstrable PowerShell administration/scripting skills.
  • Experience of Service Desk Systems.

Responsibilities

  • Provision of first-class infrastructure support to customers and systems for all Managed Service Platforms.
  • Implement observability and proactively monitor our Azure Cloud footprint, utilising dashboards, alerts and runbooks.
  • Incident response, analysis, remediation and associated post-incident root cause analysis reviews.
  • Support performance/reliability testing and capacity planning activities.
  • Perform routine system upgrades and platform patching.
  • Promote automation of repetitive tasks, reducing manual intervention.
  • Eventual participation in the 24x7 On-Call Rota.

Skills

Troubleshooting
Teamwork
Communication
Documentation
Stakeholder communication

Tools

Azure DevOps
GitHub
Terraform
PowerShell
Cloudflare

Job description

Department: TechOps
Location: Milton Keynes/Home based

Overview

The TechOps Team are responsible for the installation and support for all Technical Platforms, Operating Systems, Database Management Systems and associated products delivered to external customers of Graphnet.

The Site Reliability Engineer (SRE) will work alongside specialists within the Team, taking responsibility for all aspects of the Team’s work. With emphasis on the Technical Services function, you will be managing and supporting the infrastructure upon which Graphnet’s CareCentric product operates.

The SRE will work closely with many of Graphnet’s departments, including Operations, Security, Development, and Project Teams, ensuring that our comprehensive support service is maintained. The successful candidate will be provided with on-the-job training for all supported Platforms, Operating Systems, Databases and Applications, but is expected to have prior, demonstrable experience of Microsoft Azure, Networking and Windows Server Technologies.

  • Provision of first-class infrastructure support to customers and systems for all Managed Service Platforms
  • Implement observability and proactively monitor our Azure Cloud footprint, utilising dashboards, alerts and runbooks
  • Incident response, analysis, remediation and associated post-incident root cause analysis reviews
  • Support performance / reliability testing and capacity planning activitiesPerform routine system upgrades and platform patching
  • Promote automation of repetitive tasks, reducing manual intervention
  • Eventual participation in the 24x7 On-Call Rota
Personal Attributes
  • Excellent, demonstrable troubleshooting and problem-solving skills
  • Able to work well as an individual and as part of a teamTake part in architectural discussions with technical / non-technical team members
  • Organisational skills, with the ability to manage personal workloads in accordance with agreed timescales, whilst working under pressure
  • An eye for detail and a desire to adhere to best practices
  • Strong inter-personal and communication skills
  • Have a desire to keep up with the latest tools and techniques
  • Verbal, written communication and documentational skills
  • Ability to explain technical concepts to key stakeholders of all levels
Education and Skills
  • Minimum of four years working within an SRE function
  • Experience of service monitoring and alerting
  • Microsoft Azure PaaS components including (but not limited to):
  • Storage
  • Proven experience with Azure DevOps (ADO) and / or GitHub
  • Experience in the support of Cloudflare
  • An understanding of Terraform or other IAC tooling
  • Experience of routine infrastructure upgrades and platform patching
  • Demonstrable PowerShell administration / scripting skills
  • Experience of Service Desk Systems

Any skills or experience in the areas below would be considered an advantage for any potential candidate, but are not essential:

  • Experience or exposure of Terraform IAC software tooling
  • Understanding of Grafana Analytics and Monitoring Solution
  • Nessus Vulnerability Assessment Solution
  • Knowledge of ITIL Foundation principles and application
  • Knowledge of ISO27001, ISO27018, ISO9001 and CE+ Certifications

Your security is important to us. Graphnet Health will never request that candidates download software, share remote access to their device, or install applications in order to apply or interview for a role. As a UK-based company, we are only able to consider candidates who are currently residing in the UK and who are legally eligible to work in the UK.

Unfortunately, recruitment scams are becoming increasingly common, including fraudulent communications pretending to represent legitimate organisations. We encourage all applicants to verify communications carefully and ensure they originate from official Graphnet Health channels. If in doubt, please contact us directly before taking any action.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Graphnet Health • Milton Keynes

Hybrid
GBP 70,000 - 95,000
Azure SRE: Cloud Reliability & Observability (Remote/Hybrid)
Azure SRE: Cloud Reliability & Observability (Remote/Hybrid)

Graphnet Health Ltd. • Milton Keynes

Hybrid
GBP 70,000 - 95,000
Senior Site Reliability Engineer (LON)
Senior Site Reliability Engineer (LON)

McNally Recruitment Ltd • Greater London

Hybrid
GBP 90,000 - 150,000
Benefits as Cash
Hybrid work model
Site Reliability Engineer - NS London
Site Reliability Engineer - NS London

BAE Systems Digital Intelligence • Greater London

On-site
GBP 50,000 - 70,000
Hybrid working environment
On-call allowances
Overtime benefits for night shifts
Site Reliability Engineer – NS London
Site Reliability Engineer – NS London

BAE Systems • Greater London

On-site
GBP 45,000 - 70,000
Hybrid working flexibility
On-call allowances
Overtime benefits
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Spectrum IT Recruitment • Southampton

Hybrid
GBP 80,000 - 110,000
Senior Site Reliability Engineer (SRE)
Senior Site Reliability Engineer (SRE)

National Health Service • Greater London

Hybrid
GBP 42,000 - 52,000
Hybrid working
Flexible working arrangements
On-site core HQs
Site Reliability Engineer (remote working)
Site Reliability Engineer (remote working)

Vertus Partners • Greater London

Hybrid
GBP 77,000 - 104,000
Site Reliability Engineer
Site Reliability Engineer

Biometric Talent Ltd • Manchester

On-site
GBP 40,000 - 65,000
Performance-Based Bonus
Pension Scheme
Hybrid Working
+2
DevOps / Site Reliability Engineer (SRE)
DevOps / Site Reliability Engineer (SRE)

SCC • United Kingdom

Hybrid
GBP 70,000 - 110,000