24/7 SRE-NOC Engineer: Automate & Improve Reliability

Nice

United Kingdom

On-site

GBP 50,000 - 75,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

Nice is seeking an experienced Site Reliability Engineer (SRE – NOC) to ensure service reliability and operational excellence. The successful candidate will lead incident responses, engage in operational automation, and enhance observability. They will work closely with cross-functional teams to build monitoring solutions and improve system resilience.

Requirements include strong Linux administration skills, experience with cloud platforms, and ability to automate tasks. The role is crucial for driving service health and reliability and will involve working on major incidents and infrastructure support.

Qualifications

  • Strong Linux systems administration skills.
  • Experience with incident management and production support.
  • Prior experience in 24x7 NOC or production operations.

Responsibilities

  • Act as a primary responder in a 24x7 on-call rotation.
  • Lead Major Incident response and coordinate across teams.
  • Own service health monitoring across infrastructure and applications.
  • Automate repetitive operational tasks to reduce manual toil.
  • Develop scripts and tools to support NOC/SRE workflows.

Skills

Linux systems administration
Incident management experience
Cloud infrastructure (AWS preferred)
Containers & orchestration (Docker, Kubernetes)
Monitoring/alerting platforms
Scripting in Python, Bash, Go
Networking fundamentals (DNS, TCP/IP)

Tools

Grafana
Prometheus
Datadog
CloudWatch
Terraform
Ansible

Job description

Nice is seeking an experienced Site Reliability Engineer (SRE – NOC) to ensure service reliability and operational excellence. The successful candidate will lead incident responses, engage in operational automation, and enhance observability. They will work closely with cross-functional teams to build monitoring solutions and improve system resilience.

Requirements include strong Linux administration skills, experience with cloud platforms, and ability to automate tasks. The role is crucial for driving service health and reliability and will involve working on major incidents and infrastructure support.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SRE
SRE

Technopride Ltd • Hove

Hybrid
GBP 60,000 - 80,000
Senior Network SRE: Automation, Reliability & Observability
Senior Network SRE: Automation, Reliability & Observability

Technopride Ltd • Greater London

On-site
GBP 70,000 - 90,000
Site Reliability Engineer (DV Security Clearance)
Site Reliability Engineer (DV Security Clearance)

Onyx-Conseil • Manchester

Hybrid
GBP 90,000 - 120,000
Sr. Network Site Reliability Engineer (SREs)
Sr. Network Site Reliability Engineer (SREs)

Technopride Ltd • Greater London

On-site
GBP 70,000 - 90,000
Senior Cloud Platform SRE & Observability Lead
Senior Cloud Platform SRE & Observability Lead

NICE • Southampton

On-site
GBP 60,000 - 80,000
Senior SRE, Observability & Cloud Reliability
Senior SRE, Observability & Cloud Reliability

United States Digital Space LLC • Greater London

Hybrid
GBP 120,000 - 170,000
Hybrid work up to 3 days per week
Senior SRE: Cloud Reliability & Observability Lead
Senior SRE: Cloud Reliability & Observability Lead

Pulse Recruit • Greater London

Hybrid
GBP 65,000 - 85,000
Site Reliability Engineer
Site Reliability Engineer

Insight International (UK) Ltd • Bournemouth

On-site
GBP 55,000 - 75,000
Site Reliability Engineer - Cloud, Linux & On-Call
Site Reliability Engineer - Cloud, Linux & On-Call

Oracle • United Kingdom

On-site
GBP 70,000 - 110,000
Jira & Confluence access
Senior Network SRE: Lead Incidents & Automation
Senior Network SRE: Lead Incidents & Automation

IT WORLD LIMITED • England

On-site
GBP 70,000 - 78,000