Senior SRE: Remote, Observability & Reliability

JobCubby

Indianapolis (IN)

Hybrid

USD 115,000 - 150,000

Full time

13 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health insurance
Vision and dental insurance
401k matching
Fully remote in US

Job summary

Bloomerang is seeking an experienced Site Reliability Engineer to evolve our SRE practices, focusing on reliability, observability, and automation across services, databases, and infrastructure. You will own complex production issues, drive proactive reliability, and collaborate with Software Engineering, Product, Support, and DevOps in a fully remote role within the United States.

This is a hands-on, cross-functional role with on-call rotation and a strong emphasis on improving customer

Qualifications

  • Hands-on SRE experience applying software engineering practices to production reliability and helping establish or mature SRE practices.
  • Strong knowledge of SLIs, SLOs, error budgets, observability, automation, and toil reduction.
  • Experience building monitoring, dashboards, alerts, and telemetry using tools such as Honeycomb, New Relic, Grafana, CloudWatch, Kibana, or similar.
  • Experience with production incident management, root cause analysis, blameless post-incident reviews, and corrective-action follow-through.
  • Strong programming and scripting skills to navigate and troubleshoot application code and build automation and operational tooling.
  • Strong SQL and relational database skills for production troubleshooting; PostgreSQL experience preferred.
  • Experience troubleshooting cloud-hosted applications using source code, logs, APIs, telemetry, event streams, and databases.
  • Comfort navigating application stacks across technologies such as PHP, .NET, and Node.js.

Responsibilities

  • Own complex production support escalations and ticket triage, providing hands-on troubleshooting and resolution alongside reliability work.
  • Partner with Software Engineering to investigate complex production issues, identify root causes and reliability risks, and drive permanent solutions to recurring problems and defects.
  • Bring proven SRE practices to the team and foster proactive reliability, continuous improvement, automation, and shared ownership.
  • Lead incident response from triage and mitigation through recovery, root cause analysis, and blameless post-incident reviews, turning lessons learned into reliability improvements.
  • Build observability across products, services, and critical customer workflows using meaningful metrics, logs, traces, dashboards, and actionable alerts.
  • Define and mature SLIs and SLOs that measure system reliability and customer experience.
  • Develop synthetic monitoring for critical customer journeys to detect failures before they impact customers.
  • Identify sources of recurring operational toil and drive automation, tooling, process improvements, or permanent fixes that reduce manual effort.
  • Use AI-assisted tools and source code repositories to accelerate triage, troubleshooting, code analysis, automation, and technical investigation.
  • Participate in a rotating on-call schedule, primarily during business hours, with limited after-hours and weekend support.

Skills

SRE experience
Observability
Automation
Troubleshooting
Programming & scripting

Tools

Honeycomb
New Relic
Grafana
CloudWatch
Kibana

Job description

Bloomerang is seeking an experienced Site Reliability Engineer to evolve our SRE practices, focusing on reliability, observability, and automation across services, databases, and infrastructure. You will own complex production issues, drive proactive reliability, and collaborate with Software Engineering, Product, Support, and DevOps in a fully remote role within the United States.

This is a hands-on, cross-functional role with on-call rotation and a strong emphasis on improving customer

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Software Engineer, Site Reliability
Sr. Software Engineer, Site Reliability

JobCubby • Indianapolis (IN)

Hybrid
USD 115,000 - 150,000
Health insurance
Vision and dental insurance
401k matching
+1
Senior SRE - Hybrid, Observability & Reliability
Senior SRE - Hybrid, Observability & Reliability

Early Warning Services LLC • Chicago (IL)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Plan with match
PTO and Holidays
+1
Senior SRE: Platform Reliability & Incident Lead (Remote)
Senior SRE: Platform Reliability & Incident Lead (Remote)

Affirm, Inc. • Town of Poland (NY)

On-site
USD 32,000 - 48,000
Health insurance
Equity rewards
Flexible Spending Wallets
+1
Remote SRE — Scale, Resilience & Observability
Remote SRE — Scale, Resilience & Observability

Bright Vision Technologies • United States

On-site
USD 100,000 - 150,000
Competitive base salary
Health benefits
Long-term stability
Senior SRE: Scale Reliability, Observability & Resilience
Senior SRE: Scale Reliability, Observability & Resilience

Early Warning Services LLC • Scottsdale (AZ)

Hybrid
USD 106,000 - 130,000
Healthcare Coverage
401(k) Retirement Plan
Paid Time Off
+2
Senior SRE: Automate Reliability & Observability
Senior SRE: Automate Reliability & Observability

United States Digital Space LLC • Charlotte (TX)

On-site
USD 153,000 - 192,000
Discretionary incentive eligible
Benefits package
Senior SRE - Remote, Unlimited PTO, Reliable Platforms
Senior SRE - Remote, Unlimited PTO, Reliable Platforms

Ad Tech Industry • United States

Remote
USD 140,000 - 190,000
Unlimited PTO
Hobby & team building budget allowance
Employee Support Program
+2
Senior SRE: Reliability & Observability Lead
Senior SRE: Reliability & Observability Lead

Inspire • Atlanta (GA)

On-site
USD 140,000 - 200,000
Senior SRE: Scalable Infra, Observability & Automation
Senior SRE: Scalable Infra, Observability & Automation

Early Warning • Scottsdale (AZ)

Hybrid
USD 106,000 - 156,000
Healthcare Coverage
401(k) Company Match
Paid Time Off
+2
Senior SRE – Remote, High-Impact Cloud Infra
Senior SRE – Remote, High-Impact Cloud Infra

Upserve • United States

On-site
USD 120,000 - 180,000