Senior Platform SRE: Automate, Scale & Observe Cloud Infra

Wwshemi

United States

Remote

USD 140,000 - 190,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

BeyondTrust is seeking a Sr Site Reliability Engineer to join the Platform Infrastructure team. You will design, build, and maintain a microservices-based cloud solution using modern orchestration, automation, and IaC practices to improve reliability and scalability.

You will lead efforts around automated infrastructure, on-call incident response, and collaboration with senior engineers across teams to implement best practices.

Qualifications

  • Design long-term technical solutions and cross-team mechanisms to achieve reliability goals.
  • Define a roadmap for engineering teams to utilize automated, self-service, scalable, efficient, observable, and reliable infrastructure services as a product.
  • Align and help drive execution of the Platform Infrastructure team’s roadmap.
  • Collaborate with SREs and senior engineers across engineering organizations on best practices.
  • Provide expert technical guidance and feedback during engineering design reviews to teams onboarding to Platform Infrastructure.
  • Reduce toil (automate, automate, and automate some more).
  • Build monitoring and alerting for the Platform Infrastructure.
  • Be on an on-call rotation to respond to incidents that impact the BeyondTrust platform availability.

Responsibilities

  • Design long-term technical solutions and cross-team mechanisms to achieve reliability goals.
  • Define a roadmap for engineering teams to utilize automated, self-service, scalable, efficient, observable, and reliable infrastructure services as a product.
  • Align and help drive execution of the Platform Infrastructure team’s roadmap.
  • Collaborate with SREs and senior engineers across engineering organizations on best practices.
  • Provide expert technical guidance and feedback during engineering design reviews to teams onboarding to Platform Infrastructure.
  • Reduce toil (automate, automate, and automate some more).
  • Build monitoring and alerting for the Platform Infrastructure.
  • Be on an on-call rotation to respond to incidents that impact the BeyondTrust platform availability.

Skills

SRE best practices
Automation
Incident response
Cross-team collaboration

Tools

AWS (S3, EC2, RDS)
Kubernetes (EKS)
Istio Service Mesh
Terraform
AWS CDK
GitLab CI/CD
Datadog

Job description

BeyondTrust is seeking a Sr Site Reliability Engineer to join the Platform Infrastructure team. You will design, build, and maintain a microservices-based cloud solution using modern orchestration, automation, and IaC practices to improve reliability and scalability.

You will lead efforts around automated infrastructure, on-call incident response, and collaboration with senior engineers across teams to implement best practices.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Platform SRE: Reliability & Automation
Senior Data Platform SRE: Reliability & Automation

Empower • United States

Hybrid
USD 106,000 - 149,000
Medical, dental, vision and life保险
401(k) with company match
Tuition reimbursement
+4
Senior SRE II: Automate & Scale Cloud Reliability
Senior SRE II: Automate & Scale Cloud Reliability

Genuine Parts • Birmingham (AL)

On-site
USD 90,000 - 120,000
Senior SRE: Platform Reliability & Cloud Automation
Senior SRE: Platform Reliability & Cloud Automation

Coalition, Inc. • United States

Remote
USD 140,000 - 210,000
Medical, dental, vision coverage
Flexible PTO
Home office stipend & WeWork access
+1
Remote Platform SRE | Scale, Automate & Secure Cloud
Remote Platform SRE | Scale, Automate & Secure Cloud

First Due • United States

Remote
USD 99,000 - 165,000
Staff SRE: Scale, Observability & Automation Leader
Staff SRE: Scale, Observability & Automation Leader

Replit • Northern (KY)

Hybrid
USD 180,000 - 260,000
Salary & equity
401(k) matching
Health, dental, vision, life
+9
Senior Production SRE: Cloud & On-Prem Reliability
Senior Production SRE: Cloud & On-Prem Reliability

Weights & Biases • New York (NY)

On-site
USD 140,000 - 180,000
Medical Insurance
Dental Insurance
Vision Insurance
+15
Senior SRE: Observability, Automation & Scalable Systems
Senior SRE: Observability, Automation & Scalable Systems

Replit • Northern (KY)

Hybrid
USD 140,000 - 190,000
Competitive Salary & Equity
401(k) 4% match (US)
Health, Dental, Vision & Life
+7
Remote Senior Platform SRE — Scale Kubernetes & Observability
Remote Senior Platform SRE — Scale Kubernetes & Observability

Blue River Technology • United States

Remote
USD 148,000 - 261,000
Remote SRE / Platform Engineer — Build Reliable Cloud Platforms
Remote SRE / Platform Engineer — Build Reliable Cloud Platforms

WinTrio LLC • United States

On-site
USD 120,000 - 160,000
Healthcare
401(k)
Annual bonus
+2
Cloud Reliability Engineer: Scale, Observability, Automation
Cloud Reliability Engineer: Scale, Observability, Automation

ALVARIA • Houston (TX)

On-site
USD 120,000 - 160,000