Site Reliability Engineer - Incident & Reliability Lead

Origami Risk

United States

Hybrid

USD 100,000 - 120,000

Full time

35 hours ago
Be an early applicant
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Benefits offered by this job

Medical/Dental/Vision
401(k) match
Flexible PTO
Hybrid/Remote options
Wellness reimbursement
Life insurance
Education assistance
Pet insurance

Job summary

Origami Risk is seeking a Site Reliability Engineer to enhance time-to-resolution, reliability, and scalability of our SaaS platform. You will lead post-incident investigations, identify root causes, and implement preventive measures across client deployments.

You will configure and tune observability tooling, collaborate with Cloud Operations, SRE, and Engineering teams, and develop dashboards to proactively address performance challenges while maintaining strong incident response practices.

Qualifications

  • Bachelor's degree in Computer Science or related field or equivalent experience.
  • 5+ years of proven experience in a Site Reliability Engineering role.
  • Strong knowledge of SRE best practices and incident management protocols.
  • Deep experience using and/or configuring observability tools such as New Relic, DataDog, or Sumo Logic.
  • Proficiency in reading and writing code (JavaScript, .NET, SQL).
  • Familiarity with cloud platforms (AWS, Azure) and architectural patterns.
  • Experience operating in a Public Cloud environment (AWS preferred).

Responsibilities

  • Leads post-incident investigations for the Site Reliability team.
  • Conducts in-depth post-incident analyses to identify root causes and develop preventive strategies.
  • Drafts clear RCAs for customer delivery.
  • Cross trains colleagues on observability tools during incident and performance investigations.
  • Provides visibility to stakeholders throughout the Site Reliability process.
  • Collaborates with cross-functional teams to implement system enhancements for scalability and stability.
  • Develops client dashboards/alerts to identify performance challenges.
  • Monitors and improves time to resolution metrics.
  • Maintains observability tools for incident response and investigations.
  • Contributes to automation to streamline incident response.
  • Partners with Cloud Operations, SRE, Engineering teams and business units to advance SaaS platforms.

Skills

SRE best practices
Incident management
Observability tools
Coding: JavaScript/.NET/SQL
Cloud: AWS/Azure
CI/CD pipelines
IaC
Windows/SQL Server

Education

Bachelor's degree in Computer Science or related field

Tools

New Relic
DataDog
Sumo Logic

Job description

Origami Risk is seeking a Site Reliability Engineer to enhance time-to-resolution, reliability, and scalability of our SaaS platform. You will lead post-incident investigations, identify root causes, and implement preventive measures across client deployments.

You will configure and tune observability tooling, collaborate with Cloud Operations, SRE, and Engineering teams, and develop dashboards to proactively address performance challenges while maintaining strong incident response practices.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer - Incident & Observability Lead
Site Reliability Engineer - Incident & Observability Lead

Worky • Atlanta (GA)

Hybrid
USD 100,000 - 120,000
Medical & Dental
Hybrid/Remote work
Vision Insurance
+4
Senior Site Reliability Engineer — Scalable, Observability-Driven
Senior Site Reliability Engineer — Scalable, Observability-Driven

Origami Risk • Chicago (IL)

Hybrid
USD 100,000 - 120,000
Hybrid work arrangement
401(k) with company match
Medical, dental, and vision benefits
+1
Remote Site Reliability Engineer — Observability & Resilience
Remote Site Reliability Engineer — Observability & Resilience

Spectrum Equity • Chicago (IL)

Hybrid
USD 100,000 - 120,000
Medical and Dental coverage
401(k) with company match
Flexible PTO & holidays
+1
SRE - Incident RCA & Observability (Remote)
SRE - Incident RCA & Observability (Remote)

Socket.dev • Georgia

Hybrid
USD 100,000 - 120,000
Medical and Dental coverage
Vision coverage
401(k) match
+8
Senior Incident Command & Reliability Engineer
Senior Incident Command & Reliability Engineer

IBM • Boston (MA)

On-site
USD 140,000 - 190,000
Site Reliability Engineer
Site Reliability Engineer

Spectrum Equity • Chicago (IL)

Hybrid
USD 100,000 - 120,000
Medical and Dental coverage
401(k) with company match
Flexible PTO & holidays
+1
Site Reliability Engineer
Site Reliability Engineer

Origami Risk • Chicago (IL)

Hybrid
USD 100,000 - 120,000
Hybrid work arrangement
401(k) with company match
Medical, dental, and vision benefits
+1
Site Reliability Engineer — Incident & Deployment Expert
Site Reliability Engineer — Incident & Deployment Expert

Re Focus LLC • O’Fallon (MO)

On-site
USD 75,000 - 105,000
Site Reliability Engineer
Site Reliability Engineer

Socket.dev • Georgia

Hybrid
USD 100,000 - 120,000
Medical and Dental coverage
Vision coverage
401(k) match
+8
Site Reliability Engineer
Site Reliability Engineer

Worky • Atlanta (GA)

Hybrid
USD 100,000 - 120,000
Medical & Dental
Hybrid/Remote work
Vision Insurance
+4