Staff Field Reliability Engineer: Platform & Escalation Lead

honeycomb.io

United States

Remote

USD 200,000 - 240,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity / Stock options
Unlimited PTO
Home office stipend
Full benefits

Job summary

Honeycomb is seeking a Field Reliability Engineer to lead platform engineering, define standards for RaaS and HnyPC, and own multi-region capacity planning. You will drive instrumentation, incident response, and architecture decisions for our largest customers, working across SRE, platform, and customer teams.

You will build automation, collaborate on open source initiatives, and mentor engineers while delivering high-impact outcomes in a distributed, customer-focused environment.

Qualifications

  • 9+ years in engineering, SRE, or infrastructure with staff-level impact
  • Kubernetes (EKS strongly preferred) with production-scale deployments
  • Strong AWS across core services, multi-account architecture, cost optimization
  • Deep IaC experience: Terraform/Helm and automation tooling
  • OpenTelemetry or observability stack expertise with standards-setting experience
  • Experience in incident command and leading critical escalations
  • Proficiency in one or more languages: Go, Python, Java, TypeScript/Node.js, or .NET
  • Excellent executive communication skills with clients and leadership
  • Ability to define direction in ambiguous, high-pressure environments
  • Background leading open source contributions or ecosystem leadership

Responsibilities

  • Define architecture and standards for Refinery as a Service (RaaS) and Honeycomb Private Cloud (HnyPC) across multiple AWS accounts and regions
  • Architect Terraform modules, Helm charts, and deployment automation
  • Set technical direction for instrumentation and operations of managed infrastructure
  • Own capacity planning, scaling, upgrades, and cross-region cost optimization
  • Build tooling to scale the FRE team and improve incident response
  • Serve as final escalation point for complex customer scenarios and outages
  • Lead architecture reviews, SLO workshops, and instrumentation deep-dives for large accounts
  • Mentor and align cross-functional teams across Solutions Architecture and Engineering

Skills

Kubernetes
AWS
Incident command
Terraform
Helm
OpenTelemetry
Observability
Go
Executive communication
Ambiguity management

Tools

Terraform
Helm
Chef
Ansible

Job description

Honeycomb is seeking a Field Reliability Engineer to lead platform engineering, define standards for RaaS and HnyPC, and own multi-region capacity planning. You will drive instrumentation, incident response, and architecture decisions for our largest customers, working across SRE, platform, and customer teams.

You will build automation, collaborate on open source initiatives, and mentor engineers while delivering high-impact outcomes in a distributed, customer-focused environment.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Field Reliability Engineer (Senior)
Remote Field Reliability Engineer (Senior)

Honeycomb Enterprise • Northern (KY)

Hybrid
USD 200,000 - 240,000
Equity
Unlimited PTO
Remote-friendly
+2
Remote Senior Field Reliability Engineer - Lead Scale Infra
Remote Senior Field Reliability Engineer - Lead Scale Infra

Doist • United States

Remote
USD 200,000 - 240,000
Equity
Unlimited PTO
Remote-friendly
+3
Senior Platform Reliability Engineer - Edge & Cloud
Senior Platform Reliability Engineer - Edge & Cloud

HiveWatch • El Segundo (CA)

On-site
USD 170,000 - 190,000
Health coverage
HQ on Main Street in El Segundo
401(k) with 4% company match
+3
Senior SRE: Scale High-Traffic Systems (Remote)
Senior SRE: Scale High-Traffic Systems (Remote)

honeycomb.io • United States

Remote
USD 100,000 - 150,000
Generous equity with employee-friendly stock program
Unlimited PTO
Home office and internet stipend
+3
Staff Field Reliability Engineer
Staff Field Reliability Engineer

Honeycomb Enterprise • Northern (KY)

Hybrid
USD 200,000 - 240,000
Equity
Unlimited PTO
Remote-friendly
+2
Staff Field Reliability Engineer
Staff Field Reliability Engineer

Doist • United States

Remote
USD 200,000 - 240,000
Equity
Unlimited PTO
Remote-friendly
+3
Hybrid SRE Engineer - Cloud, Kubernetes & Automation
Hybrid SRE Engineer - Cloud, Kubernetes & Automation

Honeywell Technologies • Hamilton Township (NJ)

Hybrid
USD 99,000 - 125,000
Enterprise AI Deployment Engineer
Enterprise AI Deployment Engineer

HoneyHive • New York (NY)

On-site
USD 150,000 - 230,000
Health benefits
Vision benefits
Dental benefits
+4
Head of Field Engineering & Solutions Architecture
Head of Field Engineering & Solutions Architecture

Hone • United States

Remote
USD 180,000 - 240,000
Staff Field Reliability Engineer
Staff Field Reliability Engineer

honeycomb.io • United States

Remote
USD 200,000 - 240,000
Equity / Stock options
Unlimited PTO
Home office stipend
+1