NOC Engineer / SRE

Greenhouse Software, Inc.

United Kingdom

On-site

GBP 60,000 - 96,000

Full time

36 hours ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

NiCE in the United Kingdom is seeking an experienced SRE – NOC to join our team. You will balance traditional NOC responsibilities with reliability engineering, focusing on 24/7 service reliability, incident response, and automation.

You’ll own runbooks, design alerting, build dashboards with Grafana, and work with cross‑functional teams to reduce toil. The role suits those who engineer solutions rather than only respond to alerts, with a focus on SLOs/SLIs and scalable platforms.

Qualifications

  • Experience with incident management and production support.
  • Familiarity with cloud infrastructure (AWS preferred).
  • Monitoring or alerting platforms.
  • Scripting or programming in Python, Bash, Go, or similar.
  • Understanding of networking fundamentals (DNS, TCP/IP, load balancing).

Responsibilities

  • Act as primary or escalation responder in a 24x7 on-call rotation.
  • Lead or support Major Incident (MI) response, including triage, mitigation and resolution.
  • Coordinate across Engineering, Infrastructure, Security, and Product teams.
  • Execute and improve runbooks, playbooks and escalation paths.
  • Drive blameless post-incident reviews (PIRs) and track corrective actions.
  • Own service health monitoring across infrastructure, applications and dependencies.
  • Design and maintain alerting strategies aligned with SLIs/SLOs.
  • Build dashboards using Grafana and other monitoring tools.
  • Automate repetitive operational tasks to reduce manual toil.
  • Improve MTTD and MTTR through tooling and processes.
  • Develop scripts/tools to support NOC/SRE workflows.
  • Implement self-healing and auto-remediation where possible.
  • Collaborate with engineering to improve system design for reliability.
  • Support and troubleshoot capacity planning and availability reviews.

Skills

Incident management
24x7 NOC experience
Python
Bash
Go
Networking fundamentals
Strong communication

Tools

AWS
Grafana

Job description

At NiCE, we don’t limit our challenges. We challenge our limits. Always. We’re ambitious. We’re game changers. And we play to win. We set the highest standards and execute beyond them. And if you’re like us, we can offer you the ultimate career opportunity that will light a fire within you.

So, what's the role all about?

The SRE – NOC role sits at the intersection of traditional Network Operations Center (NOC) responsibilities and engineering‑driven reliability practices. This role focuses on 24/7 service reliability, incident response, operational automation, and observability, while actively reducing operational toil through software and automation.

Unlike a traditional NOC analyst, an SRE‑NOC is expected to engineer problems away, not just respond to alerts.

How will you make an impact?
Incident Response & Operations
  • Act as a primary or escalation responder in a 24x7 on‑call rotation
  • Lead or support Major Incident (MI) response, including triage, mitigation, and resolution
  • Coordinate across Engineering, Infrastructure, Security, and Product teams
  • Execute and improve runbooks, playbooks, and escalation paths
  • Drive blameless post‑incident reviews (PIRs) and track corrective actions
Monitoring, Alerting & Observability
  • Own service health monitoring across infrastructure, applications, and dependencies
  • Design and maintain alerting strategies that align with SLIs/SLOs
  • Build dashboards using tools such as:
  • Grafana
Reliability Engineering & Automation
  • Automate repetitive operational tasks to reduce manual toil
  • Improve mean time to detect (MTTD) and mean time to resolve (MTTR)
  • Develop scripts and tools (Python, Bash, Go, etc.) to support NOC/SRE workflows
  • Implement self‑healing and auto‑remediation where possible
  • Partner with engineering teams to improve system design for reliability
Platform & Infrastructure Support
  • Support and troubleshoot:
  • Assist with capacity planning and availability reviews
  • Ensure operational readiness for production releases
Have you got what it takes?
Technical
  • Experience with incident management and production support
  • Familiarity with:
  • Cloud infrastructure (AWS preferred)
  • Monitoring/alerting platforms
  • Scripting or programming experience in Python, Bash, Go, or similar
  • Understanding of networking fundamentals (DNS, TCP/IP, load balancing)
Operational
  • Experience working in 24x7 NOC or production operations environments
  • Ability to handle high‑pressure incidents calmly and effectively
  • Strong written and verbal communication for incident coordination
  • Comfort working from runbooks—but improving them when they fall short
Preferred / Differentiators
  • Experience defining or operating to SLOs / SLIs
  • Prior migration from traditional NOC → SRE model
  • Exposure to security, compliance, or regulated environments

Requisition ID: 11707.

Reporting into: Manager, Network Operations.

About NiCE

NICELtd. (NASDAQ: NICE)software products are used by 25,000+ global businesses, including 85 of the Fortune 100 corporations, to deliver extraordinary customer experiences,fight financial crimeand ensure public safety.Every day, NiCE software managesmore than120 million customer interactions and monitors3+billion financial transactions.

Known as an innovation powerhouse that excels in AI, cloud and digital, NiCE is consistently recognized as the market leader in its domains, with over 8,500 employees across 30+ countries.

NiCE is proud to be an equal opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, national origin, age, sex, marital status, ancestry, neurotype, physical or mental disability, veteran status, gender identity, sexual orientation or any other category protected by law.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

NOC Engineer / SRE
NOC Engineer / SRE

AI Chopping Block, Inc. • United Kingdom

Remote
GBP 65,000 - 105,000
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Greenhouse Software, Inc. • Greater London, Southampton

Hybrid
GBP 70,000 - 100,000
NICE-FLEX hybrid model
Equal opportunity employer
SRE-NOC Engineer: Master Incident Response & Reliability
SRE-NOC Engineer: Master Incident Response & Reliability

Greenhouse Software, Inc. • United Kingdom

Remote
GBP 60,000 - 96,000
Remote SRE-NOC Engineer: Automate, Observe, Resolve
Remote SRE-NOC Engineer: Automate, Observe, Resolve

AI Chopping Block, Inc. • United Kingdom

Remote
GBP 65,000 - 105,000
Technical Support Engineer
Technical Support Engineer

Nice Ltd. • United Kingdom

On-site
GBP 35,000 - 55,000
Cloud Operations Engineer - Night shift 10PM to 6AM
Cloud Operations Engineer - Night shift 10PM to 6AM

Greenhouse Software, Inc. • Greater London

Remote
GBP 60,000 - 90,000
Professional Sevices Engineer (Implementation Engineer)
Professional Sevices Engineer (Implementation Engineer)

Nice • Southampton

On-site
GBP 50,000 - 75,000
Lead Cloud Operations Engineer
Lead Cloud Operations Engineer

Greenhouse Software, Inc. • Greater London

Remote
GBP 70,000 - 95,000
Hybrid work model
Team Lead, Technical Support
Team Lead, Technical Support

Greenhouse Software, Inc. • United Kingdom

Remote
GBP 65,000 - 90,000
DevOps Engineer
DevOps Engineer

Nice • Greater London

On-site
GBP 85,000 - 110,000
NICE-FLEX hybrid model