Senior Site Reliability Engineer

Rippling, Inc.

United States

Remote

USD 150,000 - 190,000

Full time

7 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Flexible work schedule
Unlimited PTO
Medical coverage
401K
Parental leave

Job summary

NeuroFlow is seeking a Senior Site Reliability Engineer to own our production reliability across AWS, Azure, and on-prem data centers. You will define SLOs/SLIs, drive incident response, and build tooling to enable safe, rapid deployments while maintaining security and compliance.

You'll partner with engineering, security, and product to set standards, automate repetitive tasks, and document processes for longevity.

Qualifications

  • 8+ years in site reliability, DevOps, or infrastructure engineering with end-to-end ownership.
  • Built and managed AWS and Azure accounts aligned with Well-Architected Framework.
  • Deployed Docker-based software to production with reliable operation.
  • Experience across Azure, AWS, and data centers, managing Windows VMs and IIS.
  • Infrastructure-as-code with Terraform.
  • DevOps, 12-Factor App, least privilege, zero-trust in real systems.
  • Led cross-team work to define SLAs and requirements.
  • Defined and operated against SLOs, and led incident response and postmortems.

Responsibilities

  • Define SLOs and SLIs with stakeholders inside and outside engineering, back them with error budgets, and use them to decide when to ship and when to slow down.
  • Run and improve monitoring, alerting, and logging stack (Dynatrace, New Relic, Datadog, or similar) for compliance and early problem detection.
  • Take part in on-call for production incidents and help engineers resolve customer issues.
  • Lead blameless postmortems and implement root-cause fixes.
  • Automate repetitive work, measure cost savings, and report improvements.
  • Leave automation and documentation that survives ownership changes.
  • Own design, build, and upkeep of core infrastructure for rapid, safe testing and shipping.
  • Set reliability priorities and adjust with business needs.
  • Assess build vs. buy and propose proven tools to optimize engineering time.
  • Define and evolve delivery standards including CI/CD, testing, canaries, and rollback.
  • Work with security to stay compliant and raise risks before audits or incidents.
  • Mentor engineers on reliability, Azure/AWS, Terraform, and operational practices.

Skills

SRE experience
AWS
Azure
Docker
Terraform
SQL
IIS
Security compliance
DevOps
Incident response

Tools

ECS
Fargate
RDS
Lambda
S3
EventBridge
Step Functions
SQS
SNS
Dynatrace
New Relic
Datadog
Terraform

Job description

NeuroFlow CEO and West Point graduate Christopher Molaro served in the army for five years, including a tour in Iraq as a platoon leader. Coming back home, he experienced firsthand the gaps in the behavioral health system and how veterans and civilians alike face too many barriers when it comes to receiving appropriate, timely care.

While pursuing his MBA at Wharton, Chris met his future co-founder Adam Pardes, and the two agreed – even the most engaging digital mental health apps in the world wouldn’t truly change the problem; only a solution that systematically integrated behavioral health into the full healthcare ecosystem could create meaningful change. And so they created NeuroFlow.

What We Do

We pride ourselves on partnering with healthcare leaders to assist in driving better outcomes, lowering total cost of care, and making behavioral health risk more predictable and transparent. NeuroFlow exists to make sure no one who needs behavioral health support falls through the cracks.

We build more than just engaging digital health tools for self-care: we create platforms that identify population behavioral health risk early, engage individuals with acuity‑specific resources, and enable care teams to make smarter and more efficient decisions. Together, NeuroFlow’s solutions arm healthcare organizations with the insights they need to overcome the systemic challenges in today’s healthcare ecosystem.

How We Do It

The award‑winning culture at NeuroFlow is one built around encouragement and daring to be great. Our core values have been displayed in our office since day one, and each team member is responsible for carrying out these values and keeping each other accountable to them. We succeed through our flexibility and agility, navigating and transforming an industry ripe for change where "no" or "can't" is too often the default. NeuroFlow offers unique opportunities to work in a fun and challenging fast‑paced environment with direct, meaningful impact on helping to close the divide between mental and physical health.

NeuroFlow is backed by investors including HLM Venture Partners, SEMCAP Health, Concord Health Partners, Builders VC, and Dreamit, alongside grant funding from the National Science Foundation and the U.S. Department of Defense.

Our work has been recognized by TIME Magazine as one of the World's Top HealthTech Companies of 2025, by Fierce Healthcare's 2026 Fierce15, and by MedTech Breakthrough's Best Overall Mental Health Solution award.

About the Role

As our Senior Site Reliability Engineer, you'll own how reliable our production systems are and how we measure it. You'll set our SLOs, lead incident response, build the tooling that lets teams ship and experiment safely, and keep our infrastructure secure and compliant across AWS, Azure, and our data centers. You'll work with engineering, security, and product partners to set the standard for how we build and ship.

What You'll Do
Reliability & Incident Response
  • Define SLOs and SLIs with stakeholders inside and outside engineering, back them with error budgets, and use them to decide when to ship and when to slow down.
  • Run and improve our SaaS monitoring, alerting, and logging stack (Dynatrace, New Relic, Datadog, or similar) so we meet compliance requirements and catch problems before customers do.
  • Take part in on‑call for production incidents and help engineers work through customer issues.
  • Lead blameless postmortems. When a failure keeps showing up, find the root cause and fix it for good.
  • Find the manual, repetitive work that eats engineering time, measure what it costs, and automate it away. Track and report the reduction.
  • Leave behind automation and documentation that survives a change of owner.
Platform & Infrastructure
  • Own the design, build, and upkeep of the core infrastructure that lets NeuroFlow engineers test and ship product changes quickly and safely.
  • Set the priorities for reliability and platform work, bring them to your manager for approval, and adjust when business needs shift.
  • Weigh risk against impact on high‑visibility systems, and roll out improvements in steps you can measure and reverse.
  • Make the call on build vs. buy. Keep engineering time focused on technology that sets NeuroFlow apart, and recommend proven tools for the rest.
Delivery
  • Partner with engineering teams to define and evolve our delivery standards, including CI/CD, testing, and safe deployment practices like canaries and automated rollback.
Security & Compliance
  • Work with the security team to keep our infrastructure and operations compliant, and raise risks before an audit or incident surfaces them.
Influence & Mentorship
  • Be the person engineers go to for reliability, Azure, AWS, Terraform, and operational practice, and help them grow their own skills in these areas.
  • Write documentation that someone new to a system can pick up and follow.
  • Lead with solutions. When you spot a gap, come with a plan and the tradeoffs already laid out.
Qualifications
  • 8+ years in site reliability, DevOps, or infrastructure engineering, including ownership of production systems from end to end.
  • Built and managed AWS and Azure accounts and resources in line with the Well‑Architected Framework, using services such as ECS, Fargate, RDS, Lambda, SNS, SQS, S3, EventBridge, and Step Functions.
  • Deployed Docker‑based software to production, with a solid grasp of what it takes to run containers reliably.
  • Worked hands‑on across Azure, AWS, and traditional data centers, including managing Windows VMs and IIS.
  • Built deep expertise in SQL and relational database administration, including query tuning, index optimization, and resolving high‑load production incidents (SQL Server preferred).
  • Held yourself and your team to an infrastructure‑as‑code standard, with Terraform behind everything you create.
  • Put DevOps principles, the 12‑Factor App, least privilege access, and zero‑trust architecture into practice in real systems.
  • Led cross‑team work to pin down requirements and define SLAs, adjusting the technical detail for each audience.
  • Defined and operated against SLOs, and led incident response and postmortems for production outages.
  • Worked within ITIL‑aligned service management, including incident, problem, and change management.
  • Mentored other engineers on reliability and delivery practices.
Preferred Qualifications
  • Azure or AWS certifications, such as AZ‑104, AZ‑305, or AWS Certified Solutions Architect.
  • Experience with Azure Government or other federal cloud environments.
  • Experience supporting the VA, DoD, or other federal health programs, including FedRAMP or ATO work.
  • PowerShell scripting and automation.
  • Experience in a HIPAA‑compliant environment.
  • Security work beyond day‑to‑day operations, such as red teaming or penetration testing.
Security Requirements
  • Applicants selected will be subject to a security investigation and eligibility requirements for access to classified (Public Trust) information.
Company Benefits

*Applicable for full time employees

Flexible work schedule, unlimited PTO, physical and mental wellness benefits, medical coverage, parental leave, 401K, company-sponsored events, referral program, onsite gym, dog friendly office, snacks in the office, commuter benefits, onsite massages.

What We Believe

NeuroFlow prohibits unlawful discrimination against any applicant or employee on the basis of race, color, religion, gender, gender identity, gender expression, sexual orientation, national origin, family or parental status, disability, age, veteran status, or any other status protected by applicable law. All employment decisions are based on qualifications, merit, and business needs.

Applicants with disabilities may be entitled to reasonable accommodation under the terms of the Americans with Disabilities Act and certain state or local laws. A reasonable accommodation is a change in the way things are typically done which will ensure an equal employment opportunity without imposing undue hardship on NeuroFlow. Please inform our Talent team if you need any assistance completing any forms or to otherwise participate in the application process.

As a HIPAA compliant organization, NeuroFlow expects all team members to:
  • Act in accordance with NeuroFlow's Information Security Policies.
  • Protect organizational assets from unauthorized access, disclosure, modification, destruction or interference.
  • Report security events or other risks to the organization.
  • Execute organizational security processes or activities.
  • Perform security responsibilities that defined and communicated for their role.
  • Be responsible for their actions regarding the security of the organization.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

NeuroFlow • Philadelphia

On-site
USD 140,000 - 210,000
Flexible work schedule
Unlimited PTO
Medical coverage
+6
Staff .Net Engineer
Staff .Net Engineer

NeuroFlow • Philadelphia

On-site
USD 140,000 - 200,000
Flexible work schedule
Unlimited PTO
Medical coverage
+1
Senior .Net Engineer
Senior .Net Engineer

NeuroFlow • Philadelphia

On-site
USD 120,000 - 180,000
Flexible work schedule
Unlimited PTO
Medical coverage
+7
Senior Software Engineer, .Net
Senior Software Engineer, .Net

Rippling, Inc. • Philadelphia

Hybrid
USD 140,000 - 180,000
Flexible work schedule
Unlimited PTO
Medical coverage
+2
Senior Software Engineer, .Net
Senior Software Engineer, .Net

NeuroFlow • Philadelphia

On-site
USD 120,000 - 170,000
Flexible work schedule
Unlimited PTO
Medical coverage
+1
Clinical Growth Support Manager
Clinical Growth Support Manager

Rippling, Inc. • United States

Hybrid
USD 110,000 - 180,000
Flexible work schedule
Unlimited PTO
Physical and mental wellness benefits
+10
Director, Product Marketing
Director, Product Marketing

NeuroFlow • United States

On-site
USD 140,000 - 210,000
Flexible work schedule
Unlimited PTO
Medical coverage
+1
Director, Product Marketing
Director, Product Marketing

NeuroFlow • Philadelphia

On-site
USD 140,000 - 190,000
Flexible work schedule
Unlimited PTO
Medical coverage
+8
Senior Manager, Demand Generation
Senior Manager, Demand Generation

NeuroFlow • United States

Hybrid
USD 140,000 - 190,000
Flexible work schedule
Unlimited PTO
Physical and mental wellness benefits
+10
Staff Software Engineer: Growth Engineering
Staff Software Engineer: Growth Engineering

Modern Health • San Francisco (CA)

On-site
USD 180,000 - 280,000