Site Reliability Engineer

Socket.dev

Georgia

Hybrid

USD 100,000 - 120,000

Full time

7 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Medical and Dental coverage
Vision coverage
401(k) match
Flexible time off
Hybrid/Remote work
Vacation and sick leave
Wellness reimbursement
Life insurance
Education assistance
Pre-Tax Savings Accounts
EAP

Job summary

Origami Risk is hiring a Site Reliability Engineer to help improve time-to-resolution and scale our SaaS platform. You will lead post-incident investigations, craft RCAs, and collaborate across Cloud Operations and Engineering to prevent incidents and enhance observability.

The role emphasizes deep observability tool experience, cloud familiarity (AWS/Azure), and strong coding skills to support automation and incident response. Hybrid or remote work options are available where permitted.

Qualifications

  • Bachelor's degree in Computer Science or related field (or equivalent experience)
  • 5+ years of proven experience in a Site Reliability Engineering role
  • Strong knowledge of SRE best practices and incident management protocols
  • Deep experience using and/or configuring New Relic, Data Dog, SumoLogic or similar observability tools
  • Proficiency in reading and writing code (e.g., JavaScript, .NET, SQL)
  • Familiarity with cloud platforms (e.g., AWS, Azure) and architectural patterns
  • Excellent problem-solving skills and a data-driven approach to incident analysis
  • Prior experience operating within a Public Cloud environment (AWS strongly preferred)
  • Experience troubleshooting C#/.Net based web applications to identify bugs/performance challenges.
  • Solid knowledge of SaaS operations
  • Advanced written and verbal communication skills
  • Knowledge of CI/CD pipelines
  • Experience working in an IaC environment

Responsibilities

  • Leads post-incident investigations for the Site Reliability team.
  • Conducts in-depth post-incident analyses to identify root causes and develops preventive strategies.
  • Drafts clear and insightful RCAs for customer delivery.
  • Cross trains colleagues on how to best leverage observability tools during incident and performance investigations.
  • Provides visibility to all stakeholders throughout the entire Site Reliability process.
  • Collaborates with cross-functional teams to implement system enhancements that enhance scalability and stability.
  • Develops client-focused dashboards/alerts to proactively identify performance challenges.
  • Monitors and continuously improves our time to resolution metrics.
  • Maintains and configures core observability tools to ensure optimum performance and key metrics/data are available for incident response and performance investigations.
  • Provides an actionable feedback loop to Observability and Engineering teams toward improving MELT and development patterns.
  • Contributes to the development of automation tools to streamline incident response.
  • Works proactively to prevent incidents and reduce their impact on our platform.
  • Partners with the larger Cloud Operations, SRE, Engineering teams, and the business-at-large to advance our SaaS platforms.
  • Other duties as assigned.

Skills

SRE best practices
Incident management
Observability
Code reading/writing
Cloud platforms
Public cloud experience
C#/.NET troubleshooting
SaaS operations
Communication skills
CI/CD pipelines
IaC

Education

Bachelor's degree in CS or related field

Tools

New Relic
DataDog
Sumo Logic

Job description

Overview

The Site Reliability Engineer is a key force behind improving Origami’s time to resolution and advancing overall site reliability and scalability. This person participates in efforts to identify root causes during post-incident investigations, while also identifying preventative measures to minimize future disruptions. They also assist with identifying root causes in performance challenges in client implementations and implement methods for tracking key performance metrics across clients.

Starting base pay for this role is between $100,000 and $120,000. The actual base pay is dependent upon many factors, such as transferable skills, work experience, business needs, training, location, and market demands. The base pay range is subject to change and may be modified in the future. This role will be eligible for a bonus as well as competitive medical, dental, and vision benefits, wellness reimbursement, life insurance, and a 401(k) with company match. We offer vacation and sick leave benefits (under a flexible time off policy in most states).

Responsibilities
  • Leads post-incident investigations for the Site Reliability team.
  • Conducts in-depth post-incident analyses to identify root causes and develops preventive strategies.
  • Drafts clear and insightful RCAs for customer delivery.
  • Cross trains colleagues on how to best leverage observability tools during incident and performance investigations.
  • Provides visibility to all stakeholders throughout the entire Site Reliability process.
  • Collaborates with cross-functional teams to implement system enhancements that enhance scalability and stability.
  • Develops client-focused dashboards/alerts to proactively identify performance challenges.
  • Monitors and continuously improves our time to resolution metrics.
  • Maintains and configures core observability tools to ensure optimum performance and key metrics/data are available for incident response and performance investigations.
  • Provides an actionable feedback loop to Observability and Engineering teams toward improving MELT and development patterns.
  • Contributes to the development of automation tools to streamline incident response.
  • Works proactively to prevent incidents and reduce their impact on our platform.
  • Partners with the larger Cloud Operations, SRE, Engineering teams, and the business-at-large to advance our SaaS platforms.
  • Other duties as assigned.
Qualifications
  • Bachelor's degree in Computer Science or related field (or equivalent experience)
  • 5+ years of proven experience in a Site Reliability Engineering role.
  • Strong knowledge of SRE best practices and incident management protocols
  • Deep experience using and/or configuring New Relic, Data Dog, SumoLogic or similar observability tools
  • Proficiency in reading and writing code (e.g., JavaScript, .NET, SQL)
  • Familiarity with cloud platforms (e.g., AWS, Azure) and architectural patterns
  • Excellent problem-solving skills and a data-driven approach to incident analysis
  • Prior experience operating within a Public Cloud environment (AWS strongly preferred)
  • Experience troubleshooting C#/.Net based web applications to identify bugs/performance challenges.
  • Solid knowledge of SaaS operations
  • Ability to succeed when facing ambiguity and differing levels of operational maturation
  • Advanced written and verbal communication skills
  • Windows and SQL-server troubleshooting skills preferred
  • Knowledge of Continuous Integration and Continuous Delivery (CI/CD) pipelines preferred
  • Experience working in an Infrastructure as a Code (IaC) environment preferred
Benefits
  • Medical and Dental coverage available for employees, dependents, domestic partners, and spouses
  • Paid Time Off – Flexible options plus 10 paid company holidays where available**
  • All full-time positions are hybrid, with many eligible to be completely remote
  • Fully Paid by Origami Risk – Vision insurance, Short & Long-Term Disability Insurance, and Basic Life Insurance
  • Generous family leave options—including adoption and foster care placements
  • Pre-Tax Savings Accounts – Flexible Spending Account, Health Savings Account, Commuter Benefits, Dependent Care Savings Account
  • Retirement Savings – 401(k) with company match up to 4%
  • Employee Assistance Program (EAP) – Confidential & Free support offered to colleagues facing personal or work-related complications
  • Education Assistance Program – to help colleagues pursue industry/role-specific certifications
  • Wellness Benefits – reimbursement program to invest in healthy habits as well as support better colleague productivity and stress management
  • Additional coverages available – Pet Insurance, Critical Illness Insurance, and Voluntary Life & AD&D coverage

**Flexible PTO not available in California or the UK

Who We Are

Origami Risk provides integrated SaaS solutions to organizations across the risk and insurance ecosystem — from insured corporate and public entities to brokers and risk consultants, insurers, third party claims administrators (TPAs), and risk pools. We deliver our risk management and insurance core system solutions from a cloud-based platform that is highly configurable, completely scalable, and accessible via web browser and mobile app.

Dais Technology, a subsidiary of Origami Risk, provides a no-code platform that revolutionizes insurance product creation for MGAs, insurers, and reinsurers. Dais’ event-based architecture enables AI-driven bundling, automation, and real-time deployment.

Solutions from Origami Risk and Dais Technology are backed by a best-in-class service team of experienced risk and insurance professionals who possess a balance of industry knowledge and technological expertise. A singular focus on helping clients achieve their business objectives underlies our approach to developing, implementing, and supporting our risk management, safety, compliance, and insurance core system technology solutions.

Origami Risk is proud to be an equal opportunity employer. We thrive and benefit from diversity and are committed to creating an inclusive and equitable environment for all employees. We do not discriminate against any individual based upon race, religion, gender (including pregnancy, childbirth, or related medical conditions), sexual orientation, gender identity, gender expression, color, sex, national origin, age, marital status, military or veteran status, disability, or any other characteristic protected by applicable law.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer
Site Reliability Engineer

Worky • Atlanta (GA)

Hybrid
USD 100,000 - 120,000
Medical & Dental
Hybrid/Remote work
Vision Insurance
+4
Client Support Analyst II
Client Support Analyst II

Origami Risk • Seattle (WA)

Hybrid
USD 66,000 - 82,000
Medical and Dental coverage
Flexible PTO
401(k) with company match
+10
Client Support Analyst II
Client Support Analyst II

Origami Risk • Denver (CO)

Hybrid
USD 66,000 - 82,000
Medical & Dental
Flexible PTO
401(k) match
+5
AI Product Engineer
AI Product Engineer

Origami Risk • United States

Hybrid
USD 117,000 - 145,000
Medical & Dental
Vision Insurance
401(k) with company match
+5
Regional Sales Manager
Regional Sales Manager

Origami Risk • Chicago (IL)

On-site
USD 165,000 - 185,000
Medical and Dental coverage
401(k) with company match
Hybrid work environment
+1
(Senior) Sales Executive, Integrated Risk Management
(Senior) Sales Executive, Integrated Risk Management

Origami Risk • United States

Hybrid
USD 130,000 - 160,000
Medical + Dental
PTO + holidays
Hybrid/Remote options
+8
Product Manager (GRC)
Product Manager (GRC)

Origami Risk • United States

Hybrid
USD 120,000 - 145,000
Medical and Dental coverage
Paid Time Off
Hybrid work options
+5
Product Manager (RMIS)
Product Manager (RMIS)

Origami Risk • United States

Hybrid
USD 120,000 - 145,000
Medical and Dental coverage
401(k) with company match
Flexible PTO
Full Stack Engineer at Origami Risk
Full Stack Engineer at Origami Risk

Feedinkoo • United States

Hybrid
USD 100,000 - 122,000
Medical and Dental coverage
Vision insurance
401(k) with company match
+9
Client Success Strategy & Operations Lead
Client Success Strategy & Operations Lead

Origami Risk • Atlanta (GA)

Hybrid
USD 128,000 - 160,000
Medical & Dental
Vision Insurance
Flexible PTO
+6