Senior Observability Engineer: Build Reliability & Dashboards

Ripple

New York (NY)

On-site

USD 160,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Ripple in New York, NY is seeking a Senior Site Reliability Engineer, Observability, to lead hands‑on observability work while coaching product teams toward higher reliability. You will design monitoring, set SLOs/SLIs, reduce alert noise, and optimize observability costs across Azure and AWS, with Terraform as a core tool.

You will also own incident management foundations, including on‑call rotations, postmortems, and runbooks, while partnering with cross‑functional teams to improve MTTR/MTTD

Qualifications

  • 7+ years in Site Reliability Engineering, DevOps, or Platform Engineering with a strong focus on observability and production operations.
  • Proven ability to deliver hands‑on engineering work while coaching and mentoring teams—comfortable switching between builder and consultant modes.
  • Experience working in Agile/Scrum environments and collaborating effectively with cross‑functional teams.

Responsibilities

  • Design and implement monitoring, alerting, and dashboards in New Relic across Azure and AWS; write NRQL queries for troubleshooting and reporting.
  • Define and implement SLOs/SLIs and error budgets; coach teams on using them to balance feature velocity with reliability.
  • Lead alert noise reduction and signal quality engineering—tune thresholds, eliminate false positives, and ensure every alert is actionable.
  • Develop Terraform infrastructure as code for provisioning and managing monitoring resources and observability infrastructure with governance.
  • Author and troubleshoot Azure DevOps pipelines; support deployment visibility, change tracking, and release hygiene related to production reliability.
  • Administer Incident.IO: alert routing, notification workflows, Slack and OpsGenie integration, runbook management.
  • Build incident management foundations: PIR/postmortem processes, on‑call rotation design, escalation policies, and incident severity classifications.
  • Track MTTR, MTTD, and incident frequency; drive continuous improvement with engineering teams.
  • Respond to and debrief on production incidents with real‑time troubleshooting and structured post‑incident reviews.

Skills

New Relic NRQL
Structured logging
Distributed tracing
SLOs/SLIs
Incident management tooling
Postmortem reviews

Tools

Terraform
PowerShell
Azure
AWS
Azure DevOps
Octopus Deploy
Slack
Jira

Job description

Ripple in New York, NY is seeking a Senior Site Reliability Engineer, Observability, to lead hands‑on observability work while coaching product teams toward higher reliability. You will design monitoring, set SLOs/SLIs, reduce alert noise, and optimize observability costs across Azure and AWS, with Terraform as a core tool.

You will also own incident management foundations, including on‑call rotations, postmortems, and runbooks, while partnering with cross‑functional teams to improve MTTR/MTTD

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Observability & Reliability Engineer
Senior Observability & Reliability Engineer

Ripple • New York (NY)

On-site
USD 160,000 - 200,000
Competitive salary
Equity
Wellness benefits
+2
Senior Observability & Reliability Engineer
Senior Observability & Reliability Engineer

blockchaincapital.com • New York (NY)

On-site
USD 130,000 - 180,000
Senior Observability & SRE Engineer
Senior Observability & SRE Engineer

Hidden Road • Chicago (IL)

Hybrid
USD 160,000 - 200,000
Equity
Bonuses
Healthcare
+4
Senior SRE: Observability & Incident Management
Senior SRE: Observability & Incident Management

Hidden Road • New York (NY)

Hybrid
USD 160,000 - 200,000
Equity
Bonuses
Healthcare
+4
Senior SRE: Observability & Incident Champion (Azure/AWS)
Senior SRE: Observability & Incident Champion (Azure/AWS)

Hidden Road • New York (NY)

Hybrid
USD 160,000 - 200,000
Senior SRE & Observability Engineer
Senior SRE & Observability Engineer

Hidden Road • Chicago (IL)

On-site
USD 160,000 - 200,000
Senior Site Reliability Engineer, Observability
Senior Site Reliability Engineer, Observability

blockchaincapital.com • New York (NY)

On-site
USD 130,000 - 180,000
Senior Site Reliability Engineer - Azure & Observability
Senior Site Reliability Engineer - Azure & Observability

Luxoft • United States

On-site
USD 140,000 - 180,000
Senior Site Reliability Engineer, Observability New York, NY, United States
Senior Site Reliability Engineer, Observability New York, NY, United States

Ripple • New York (NY)

On-site
USD 160,000 - 200,000
Senior SRE: Observability, CI/CD & Security
Senior SRE: Observability, CI/CD & Security

Ripple • Chicago (IL)

On-site
USD 160,000 - 200,000
Competitive salary
Comprehensive healthcare benefits
Employee giving match
+2