Senior SRE & Backend Engineer - Reliability at Scale(Remote)

Nabla

Buffalo (NY)

Hybrid

USD 140,000 - 210,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary
Stock options
Medical coverage
Unlimited PTO
Sick leave
Parental leave
Home office grant
Flexible schedule

Job summary

Nabla is seeking a Senior SRE/Platform Engineer to own production reliability and scale our on-call culture across a fast-paced, production-critical platform.

You will instrument every layer with monitoring and observability, drive the SRE roadmap from SLIs to chaos engineering, and collaborate with backend, ML, and frontend squads to embed reliability into deployments. This role emphasizes autonomy, strong coding fundamentals, and effective communication during incidents.

Qualifications

  • Senior-level SRE or infrastructure experience in production-critical environments.
  • Hands-on with cloud infrastructure (GCP preferred; AWS/Azure acceptable).
  • Structured on-call with alerting, runbooks, post-mortems.
  • Software engineering fundamentals: IaC, Terraform PRs, debugging incidents.
  • Autonomy and ability to drive priorities.
  • Clear communication with engineers and leadership, even during incidents.
  • Bonus: Kubernetes, GCP tooling, PostgreSQL at scale, or healthcare security/compliance.

Responsibilities

  • Own production reliability across the full stack.
  • Own and scale the on-call culture with runbooks and escalation.
  • Instrument monitoring, alerting, and observability to surface issues.
  • Drive the SRE roadmap with leadership on reliability investments.
  • Make engineers faster via tooling and reduced toil.
  • Collaborate across squads with back-end, ML, and front-end engineers.

Skills

SRE experience
Infrastructure
Backend engineering
On-call experience
Cloud familiarity
Terraform PRs
Communication

Tools

Kubernetes

Job description

Nabla is seeking a Senior SRE/Platform Engineer to own production reliability and scale our on-call culture across a fast-paced, production-critical platform.

You will instrument every layer with monitoring and observability, drive the SRE roadmap from SLIs to chaos engineering, and collaborate with backend, ML, and frontend squads to embed reliability into deployments. This role emphasizes autonomy, strong coding fundamentals, and effective communication during incidents.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior SRE: Build Scalable, Reliable Platforms — Remote
Senior SRE: Build Scalable, Reliable Platforms — Remote

Nord Security • Town of Poland (NY)

On-site
USD 140,000 - 200,000
Premium healthcare
Work from anywhere
Mentorship programs
+3
Senior SRE: Platform Reliability & Incident Lead (Remote)
Senior SRE: Platform Reliability & Incident Lead (Remote)

Affirm, Inc. • Town of Poland (NY)

On-site
USD 32,000 - 48,000
Health insurance
Equity rewards
Flexible Spending Wallets
+1
Senior Platform SRE: Scale & Reliability
Senior Platform SRE: Scale & Reliability

United States Digital Space LLC • United States

Hybrid
USD 103,000 - 162,000
Health insurance
Vacation and RTT
Mental health and coaching
+7
SRE / Backend Engineer
SRE / Backend Engineer

Nabla • Buffalo (NY)

Hybrid
USD 140,000 - 210,000
Competitive salary
Stock options
Medical coverage
+5
Senior SRE & Software Engineer: Domain Reliability Lead
Senior SRE & Software Engineer: Domain Reliability Lead

Hispanic Alliance for Career Enhancement • Richardson (TX)

On-site
USD 93,000 - 204,000
Senior SRE - Cloud & Observability
Senior SRE - Cloud & Observability

Ridgeline • Reno (NV)

Hybrid
USD 153,000 - 210,000
Unlimited vacation
Education reimbursement
Wellness reimbursement
+1
Senior SRE: Scale Reliability Leader (Hybrid)
Senior SRE: Scale Reliability Leader (Hybrid)

Plenful • San Francisco (CA)

Hybrid
USD 180,000 - 250,000
Healthcare Coverage
401(k) with Company Match
Equity
+5
Senior SRE: Scale Reliability & Observability
Senior SRE: Scale Reliability & Observability

Megaport • Abbeyville (CO)

On-site
USD 130,000 - 190,000
Contractor (PJ)
Paid Time Off
Competitive Compensation
+4
Senior SRE: Remote Platform Reliability & Automation
Senior SRE: Remote Platform Reliability & Automation

Halo Media • United States

Remote
USD 140,000 - 200,000
Remote Observability & Reliability Engineer
Remote Observability & Reliability Engineer

Nscale • United States

Hybrid
USD 145,000 - 180,000
Medical, dental, vision
Flexible paid time off
Parental leave
+1