Senior Site Reliability Engineer

BetterUp

New York (NY)

Hybrid

USD 164,000 - 205,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Access to BetterUp coaching
Competitive compensation plan
Medical, dental, and vision insurance
Flexible paid time off
Learning and Development stipend
401(k) self contribution

Job summary

A leading technology firm is seeking an experienced Site Reliability Engineer to transform production systems using AI tools and maintain cloud infrastructure on AWS. Candidates should have over four years of SRE experience, strong skills in Kubernetes, and a genuine excitement for AI tooling. Join a hybrid work environment with a focus on collaboration and enjoy competitive compensation, benefits, and opportunities for advancement.

Qualifications

  • 4+ years of experience in SRE or infrastructure roles.
  • Hands-on Kubernetes experience: deploying, scaling, debugging, and securing clusters.
  • Genuine excitement about AI tooling.

Responsibilities

  • Leverage AI-powered tools to monitor and maintain production systems.
  • Build and operate cloud infrastructure on AWS using Terraform.
  • Manage and scale Kubernetes clusters.

Skills

SRE experience
AWS expertise
Kubernetes management
Terraform skills
Observability knowledge
Strong debugging skills
Clear communication
Automation mindset

Tools

AWS
Terraform
Kubernetes
Datadog
Prometheus
OpenTelemetry

Job description

Get AI-powered advice on this job and more exclusive features.

Let’s face it, a company whose mission is human transformation better have some fresh thinking about the employer/employee relationship. We do. We can’t cram it all in here, but you’ll start noticing it from the first interview. Even our candidate experience is different. And when you get an offer from us (and accept it), you get way more than a paycheck. You get a personal BetterUp Coach, a development plan, a trained and coached manager, the most amazing team you’ve ever met (yes, each with their own personal BetterUp Coach), and most importantly, work that matters. This makes for a remarkably focused and fulfilling work experience. Frankly, it’s not for everyone. But for people with fire in their belly, it’s a game-changing, career-defining, soul-lifting move.

Join us and we promise you the most intense and fulfilling years of your career, doing life-changing work in a fun, inventive, soulful culture. If that sounds exciting—and the job description below feels like a fit—we really should start talking.

We are a hybrid company with a focus on in-person collaboration when necessary. Employees are expected to be available to work from one of our office hubs at least two days per week, or eight days per month. Our US hub locations include: Austin, TX; Chicago, IL; New York City, NY; San Francisco, CA; and the Washington, DC metro area. If this is a role based in Europe, our Europe hub locations are London, UK and Amsterdam, NL. Please ensure you can realistically commit to this structure before applying.

What You’ll Do
  • Leverage AI-powered tools and automation to transform how we monitor, troubleshoot, and maintain production systems
  • Build and operate cloud infrastructure on AWS, using Terraform to codify and version-control our entire environment
  • Manage and scale Kubernetes clusters that power BetterUp's platform, ensuring high availability and performance
  • Design intelligent alerting and observability systems
  • Collaborate with engineering teams to embed reliability into the development lifecycle, shifting left on operational concerns
  • Automate incident response workflows and build self-healing infrastructure
  • Experiment with and adopt emerging AI tools for log analysis, anomaly detection, and predictive maintenance
  • Drive continuous improvement through data-driven retrospectives and reliability metrics
If you have some or all of the following, please apply
  • 4+ years of experience in SRE or infrastructure roles
  • Genuine excitement about AI tooling: you're already using copilots, AI assistants, or LLM-based tools in your workflow and are excited to push your skillset further in this area
  • Deep experience with AWS
  • Hands‑on Kubernetes experience: deploying, scaling, debugging, and securing clusters
  • Strong Terraform skills with experience managing complex, multi-environment infrastructure
  • Familiarity with modern observability stacks (Datadog, Prometheus, OpenTelemetry)
  • Strong debugging instincts and comfort navigating distributed systems
  • Clear communication skills - you can explain a production incident to engineers and executives alike
  • A builder's mindset: you see manual processes as opportunities for automation
AI at BetterUp

Our team thrives at the intersection of human expertise and AI capability. As an AI-forward company, adaptation and continuous learning are part of our daily work. We’re looking for teammates who are excited to evolve alongside technology – people who experiment boldly, share their discoveries openly, and help define best practices for AI-augmented work. These professionals thoughtfully integrate AI into their work to deliver exceptional results while maintaining the human judgment and creativity that drives real innovation. During our interview process, you’ll have opportunities to showcase how you harness AI to learn, iterate, and amplify your impact.

Benefits
  • Access to BetterUp coaching; one for you and one for a friend or family member
  • A competitive compensation plan with opportunity for advancement
  • Medical, dental, and vision insurance
  • Flexible paid time off
  • Per year:
    • All federal/statutory holidays observed
    • 4 BetterUp Inner Workdays (https://www.betterup.co/inner-work)
    • 5 Volunteer Days to give back
    • Learning and Development stipend
    • Company wide Summer & Winter breaks
  • Year‑round charitable contribution of your choice on behalf of BetterUp
  • 401(k) self contribution

We are dedicated to building diverse teams that fuel an authentic workplace and sense of belonging for each and every employee. We know applying for a job can be intimidating, please don’t hesitate to reach out — we encourage everyone interested in joining us to apply.

BetterUp Inc. provides equal employment opportunities (EEO) to all employees and applicants for employment without regard to race, color, religion, sex, national origin, disability, genetics, gender, sexual orientation, age, marital status, veteran status. In addition to federal law requirements, BetterUp Inc. complies with applicable state and local laws governing nondiscrimination in employment in every location in which the company has facilities. This policy applies to all terms and conditions of employment, including recruiting, hiring, placement, promotion, termination, layoff, recall, transfer, leaves of absence, compensation, and training.

At BetterUp, we compensate our employees fairly for their work. Base salary is determined by job‑related experience, education/training, residence location, as well as market indicators. The range below is representative of base salary only and does not include equity, sales bonus plans (when applicable) and benefits. This range may be modified in the future. The base salary range for the role is as follows:
New York and San Francisco: $164,000 – $205,000
Austin and Arlington (D.C. Area): $147,600 – $184,500

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

BetterUp • Austin (TX)

Hybrid
USD 147,000 - 185,000
Access to BetterUp coaching
Medical, dental, and vision insurance
Flexible paid time off
+2
Senior Data Platform Engineer for AI & Analytics
Senior Data Platform Engineer for AI & Analytics

BetterUp • Austin (TX), New York (NY), San Francisco (CA), Arlington (VA)

Hybrid
GBP 145,000 - 182,000
BetterUp coaching
Health insurance
PTO (flexible)
+3
Staff Full-Stack Engineer — Customer Value + Performance Intelligence (Tech Lead)
Staff Full-Stack Engineer — Customer Value + Performance Intelligence (Tech Lead)

BetterUp • Austin (TX)

On-site
USD 194,000 - 243,000
Access to BetterUp coaching
Medical, dental, and vision insurance
Flexible paid time off
+2
Staff Full-Stack Engineer — Customer Value + Performance Intelligence (Tech Lead)
Staff Full-Stack Engineer — Customer Value + Performance Intelligence (Tech Lead)

betterup • Bowlin (WV)

Hybrid
USD 194,000 - 243,000
Access to BetterUp coaching
401(k) self contribution
Flexible paid time off
Strategic AI Product Lead for Enterprise Solutions
Strategic AI Product Lead for Enterprise Solutions

BetterUp • New York (NY)

Hybrid
USD 236,000 - 296,000
Client Delivery Director
Client Delivery Director

BetterUp • New York (NY)

Hybrid
USD 215,000 - 250,000
Access to BetterUp coaching for you and a friend
Competitive compensation plan
Medical, dental, and vision insurance
+4
Staff Data Platform Engineer
Staff Data Platform Engineer

BetterUp • Austin (TX), New York (NY), San Francisco (CA), Arlington (VA)

Hybrid
GBP 145,000 - 182,000
BetterUp coaching
Health insurance
PTO (flexible)
+3
Staff Full-Stack Engineer — Customer Value + Performance Intelligence (Tech Lead)
Staff Full-Stack Engineer — Customer Value + Performance Intelligence (Tech Lead)

BetterUp • San Francisco (CA)

On-site
USD 216,000 - 270,000
Access to BetterUp coaching
Medical, dental, and vision insurance
Flexible paid time off
+2
Staff Solutions Engineer
Staff Solutions Engineer

Motive Software • San Francisco (CA)

Hybrid
USD 165,000 - 207,000
Access to BetterUp coaching
Competitive compensation plan
Medical, dental and vision insurance
+3
Senior Solutions Engineer
Senior Solutions Engineer

BetterUp • Austin (TX)

On-site
USD 129,000 - 178,000
Access to BetterUp coaching
Medical, dental and vision insurance
Flexible paid time off
+2