Site Reliability Engineer II: Automation & Reliability

Webhosting

Northern (KY)

Hybrid

USD 110,000 - 160,000

Full time

42 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Backblaze, the object storage leader, is seeking a Site Reliability Engineer II to help ensure the stability, scalability, and reliability of our services and infrastructure. You will focus on building automation, maintaining observability, and supporting incident response to keep customer-facing systems performing at their best.

The SRE will collaborate with engineering, product, and operations teams to embed reliability practices into day-to-day development and operations while contributing to

Responsibilities

  • Support the availability and durability of critical services across production environments.
  • Monitor service health using SLIs, SLOs, and error budgets, and elevate issues when thresholds are at risk.
  • Participate in on-call rotations, incident response, and post-incident reviews to drive service improvements.
  • Follow established ITIL/OSS processes (incident, change, problem, and capacity management).
  • Develop automation for common operational tasks, reducing manual intervention and toil.
  • Contribute to monitoring, logging, and alerting frameworks (e.g., Prometheus, Grafana, Catchpoint, ELK).
  • Work with CI/CD pipelines, configuration management, and infrastructure as code tools (Terraform, Ansible, Jenkins).
  • Write scripts (Bash, Python, Go, etc.) to improve system reliability and efficiency.
  • Partner with engineering, product, and operations teams to support resilient system design and operations.
  • Assist in capacity planning and disaster recovery exercises.
  • Work with vendors and service providers to troubleshoot service issues and track SLA performance.
  • Document systems, share learnings, and help grow a reliability-minded engineering culture.
  • Contribute to playbooks, runbooks, and operational documentation.
  • Identify recurring issues and propose long-term improvements.
  • Promote reliability-focused practices within development and operations teams.

Job description

Backblaze, the object storage leader, is seeking a Site Reliability Engineer II to help ensure the stability, scalability, and reliability of our services and infrastructure. You will focus on building automation, maintaining observability, and supporting incident response to keep customer-facing systems performing at their best.

The SRE will collaborate with engineering, product, and operations teams to embed reliability practices into day-to-day development and operations while contributing to

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Site Reliability Engineer II
Site Reliability Engineer II

Webhosting • Northern (KY)

Hybrid
USD 110,000 - 160,000
Site Reliability Engineer: Build Resilient, Automated Cloud
Site Reliability Engineer: Build Resilient, Automated Cloud

SRE • Puerto Rico

Hybrid
USD 120,000 - 180,000
Senior Site Reliability Engineer – Remote, Impact & Automation
Senior Site Reliability Engineer – Remote, Impact & Automation

Midwest Startups • United States

On-site
USD 175,000 - 185,000
Market-leading medical, dental, and視on
Stock options
Premium-Tier Origin Financial Wellness
+6
Site Reliability Engineer II: Scale, Automation & Observability
Site Reliability Engineer II: Scale, Automation & Observability

5014 Disney Entertainment & Sports LLC • New York (NY)

On-site
USD 123,000 - 165,000
Site Reliability Engineer II: Cloud, Automation & Observability
Site Reliability Engineer II: Cloud, Automation & Observability

Restaurant365 • San Francisco (CA)

Hybrid
USD 98,000 - 139,000
Site Reliability Engineer II - Automate, Scale, Resilience
Site Reliability Engineer II - Automate, Scale, Resilience

Motion • Birmingham (AL)

On-site
USD 80,000 - 110,000
Healthcare coverage
401(k)
Tuition reimbursement
+3
Lead Site Reliability Engineer: AWS Cloud & Automation
Lead Site Reliability Engineer: AWS Cloud & Automation

Selby Jennings • Wilmington (NC)

On-site
USD 140,000 - 200,000
Senior Site Reliability Engineer - Cloud Platform & Automation
Senior Site Reliability Engineer - Cloud Platform & Automation

Ridgeline, Inc. • San Ramon (CA)

Hybrid
USD 153,000 - 210,000
Unlimited vacation
Educational reimbursement
Comprehensive insurance plans
Senior Site Reliability Engineer: Reliability at Scale
Senior Site Reliability Engineer: Reliability at Scale

Veloc Inc • Coppell (TX)

On-site
USD 140,000 - 190,000
Senior SRE II: Scale, Reliability & Automation
Senior SRE II: Scale, Reliability & Automation

LexisNexis Risk Solutions • Northern (KY)

Hybrid
USD 105,000 - 175,000