Senior Site Reliability Engineer - Hybrid & Observability

Talentify

Woonsocket (RI)

Hybrid

USD 150,000 - 190,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health benefits
Referral program
Growth opportunities

Job summary

Talentify is seeking an experienced Site Reliability Engineer to own end-to-end reliability of critical platforms across hybrid cloud and on-prem environments in Woonsocket, RI. You will drive SLI/SLO health, incident response, and observability, collaborating with product and operations to embed reliability in design and delivery.

You will lead postmortems, perform root cause analysis, and mentor teams in chaos engineering, fault injection, and SRE best practices, shaping organizational

Qualifications

  • 8+ years of Senior Software engineering experience in SRE, DevOps, platform engineering, or related production-systems roles in distributed systems at production scale with active on-call responsibility.
  • Experience tuning and validating timeseries anomaly detection models in a production observability context.
  • Strong programming proficiency in Python, React, and Java at production quality.
  • Hands-on experience designing SLIs, SLOs, and managing error budgets for customer-facing or business-critical services.

Responsibilities

  • Own and drive the end-to-end reliability, availability, and performance of critical retail and pharmacy technology platforms across hybrid cloud and on-premises environments.
  • Establish and maintain SLI/SLO health, alerting strategies, observability standards, and business-aligned monitoring for the assigned application domain.
  • Lead production incident response as Incident Commander, drive root cause analysis, postmortems, and continuous reliability improvements.
  • Partner with engineering, product, and operations teams to embed reliability, resiliency, scalability, and operational readiness into system design and delivery.
  • Build and optimize automation, self-service capabilities, and operational tooling to eliminate toil, improve efficiency, and reduce manual intervention.
  • Design and execute proactive reliability initiatives, including production readiness reviews, dependency risk assessments, fault injection, and chaos engineering exercises.
  • Mentor engineers, champion SRE best practices, and enable teams to independently detect, respond to, and learn from production issues with minimal SRE involvement.

Skills

Python
React
Java
Incident Commander
Observability
Anomaly detection
SRE
DevOps
Chaos engineering
LLM integration
Kafka
Istio
Envoy
Terraform
Ansible
Kubernetes
GCP
Rancher K3s
Airflow
Tidal
BigQuery
PostgreSQL

Tools

Prometheus
Grafana
Open Telemetry
Loki
Splunk
Elasticsearch

Job description

Talentify is seeking an experienced Site Reliability Engineer to own end-to-end reliability of critical platforms across hybrid cloud and on-prem environments in Woonsocket, RI. You will drive SLI/SLO health, incident response, and observability, collaborating with product and operations to embed reliability in design and delivery.

You will lead postmortems, perform root cause analysis, and mentor teams in chaos engineering, fault injection, and SRE best practices, shaping organizational

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior SRE - Hybrid, AWS & Observability
Senior SRE - Hybrid, AWS & Observability

Early Warning • San Francisco (CA)

Hybrid
USD 128,000 - 156,000
Healthcare coverage
401(k) plan
Paid time off
+2
Senior SRE - Hybrid, Observability & Reliability
Senior SRE - Hybrid, Observability & Reliability

Early Warning Services LLC • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Plan with match
PTO and Holidays
+1
Senior Site Reliability Engineer - Hybrid, High Impact
Senior Site Reliability Engineer - Hybrid, High Impact

Early Warning • Scottsdale (AZ)

Hybrid
USD 106,000 - 156,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+2
Site Reliability Engineer
Site Reliability Engineer

Talentify • Woonsocket (RI)

Hybrid
USD 150,000 - 190,000
Health benefits
Referral program
Growth opportunities
Principal Site Reliability Engineer – Hybrid Role
Principal Site Reliability Engineer – Hybrid Role

Early Warning Services LLC • San Francisco (CA)

Hybrid
USD 194,000 - 284,000
Healthcare coverage
401(k) with company match
Paid time off
Senior Site Reliability Engineer - Reliability Leader
Senior Site Reliability Engineer - Reliability Leader

Early Warning Services LLC • San Francisco (CA)

Hybrid
USD 128,000 - 156,000
Healthcare Coverage
401(k) Match
Paid Time Off
+2
Senior Staff SRE — Reliability, Observability & Automation
Senior Staff SRE — Reliability, Observability & Automation

Early Warning Services LLC • Scottsdale (AZ)

Hybrid
USD 150,000 - 200,000
Healthcare coverage
401(k) with company match
Paid time off & holidays
+2
Hybrid Principal SRE - Reliability at Scale & Observability
Hybrid Principal SRE - Reliability at Scale & Observability

Early Warning Services LLC • Scottsdale (AZ)

Hybrid
USD 194,000 - 237,000
Healthcare Coverage
401(k) Plan
Paid Time Off
+2
Senior Site Reliability Engineer — Remote, AWS & Observability
Senior Site Reliability Engineer — Remote, AWS & Observability

Prove • United States

Hybrid
USD 140,000 - 190,000
Wellbeing reimbursement
401k Match
Parental Leave Policy
+5
Senior SRE: Observability & Automation Leader (Hybrid)
Senior SRE: Observability & Automation Leader (Hybrid)

Talentify • Holyoke (MA)

Hybrid
USD 134,000 - 170,000
Hybrid work environment
Relocation assistance
Tuition reimbursement
+6