Senior Observability SRE — Global Monitoring (Hybrid)

Nord Security

Town of Poland (NY)

Hybrid

USD 140,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work (3 office days, 2 remote)
Growth programs
Private health insurance
Wellness programs
Company events & team-building
Work from anywhere

Job summary

Nord Security is seeking a Senior Site Reliability Engineer (SRE) focused on observability to design monitoring systems, reduce alert fatigue, and partner with data teams on anomaly detection. You will own health signals across our distributed infrastructure.

As part of a global team, you will influence how we monitor, alert, and optimize performance across services, ensuring reliable delivery for millions of users.

Qualifications

  • Experience designing monitoring for distributed systems.
  • Ability to design golden signals and actionable alerts.
  • Proficiency in Python scripting.
  • Strong Linux administration and debugging.
  • Solid networking fundamentals.
  • Bonus: SaltStack and advanced networking knowledge.

Responsibilities

  • Design, build, and improve monitoring pipelines and observability tooling across globally distributed infrastructure
  • Define and implement service-level monitoring based on golden signals (latency, traffic, errors, saturation)
  • Reduce alert fatigue - build meaningful, actionable alerts that engineers trust
  • Develop and maintain custom exporters, scripts, and integrations for metrics and log collection
  • Collaborate with the data team on anomaly detection and data-driven operational insights
  • Understand service signals - know what to measure, why, and what the numbers actually mean

Skills

Python
Linux
Networking fundamentals
Observability
Monitoring pipelines
On-call management

Tools

Naemon (Nagios)
Prometheus exporters
Telegraf
Fluent Bit
VictoriaMetrics
OpenSearch
Grafana

Job description

Nord Security is seeking a Senior Site Reliability Engineer (SRE) focused on observability to design monitoring systems, reduce alert fatigue, and partner with data teams on anomaly detection. You will own health signals across our distributed infrastructure.

As part of a global team, you will influence how we monitor, alert, and optimize performance across services, ensuring reliable delivery for millions of users.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE, Observability & Reliability Platform
Senior SRE, Observability & Reliability Platform

Chainlink Labs • United States

On-site
USD 120,000 - 160,000
Flexible working hours
Career growth opportunities
Global remote team
+1
Senior Observability Platform SRE — Scale & Automate
Senior Observability Platform SRE — Scale & Automate

Dimensional Fund Advisors • Charlotte (NC)

Hybrid
USD 120,000 - 150,000
Comprehensive benefits
Educational initiatives
Celebrations of history and culture
Senior SRE - Cloud & Observability
Senior SRE - Cloud & Observability

Ridgeline • Reno (NV)

Hybrid
USD 153,000 - 210,000
Unlimited vacation
Education reimbursement
Wellness reimbursement
+1
Staff SRE — Global Platform & Infra Lead
Staff SRE — Global Platform & Infra Lead

Nord Security • Town of Poland (NY)

Hybrid
USD 180,000 - 230,000
Hybrid work
Premium healthcare
Work from anywhere
+3
Senior Site Reliability Engineer - Build SRE Ops (Hybrid)
Senior Site Reliability Engineer - Build SRE Ops (Hybrid)

Mission Staffing • New York (NY)

Hybrid
USD 80,000 - 100,000
Senior SRE & Platform Engineer — Observability & Automation
Senior SRE & Platform Engineer — Observability & Automation

Techunting • United States

On-site
USD 120,000 - 150,000
Senior SRE - Observability & Performance Eng (Hybrid)
Senior SRE - Observability & Performance Eng (Hybrid)

ManpowerGroup Global, Inc. • Town of Norway (WI), Chandler (AZ)

Hybrid
USD <1,000
Medical and Prescription Drug Plans
Dental Plan
Vision Plan
+2
Lead Observability & SRE Architect
Lead Observability & SRE Architect

CRC Group • Charlotte (NC)

On-site
USD 130,000 - 180,000
Senior Engineer - Site Reliability Engineering
Senior Engineer - Site Reliability Engineering

LSEG (London Stock Exchange Group) • Allen (TX)

On-site
USD 140,000 - 190,000
Observability SRE — Incident Response & Dashboards
Observability SRE — Incident Response & Dashboards

United States Digital Space LLC • San Francisco (CA)

Hybrid
USD 120,000 - 150,000