Senior Site Reliability Engineer

Socket.dev

New York (NY)

Hybrid

USD 200,000 - 240,000

Full time

2 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Hybrid schedule 3 days/week in NYC
Equity in growing company
Health/dental/vision + 401k
Paid time off and parental leave
Autonomy to ship meaningful work

Job summary

Kontakt.io is seeking a Senior Site Reliability Engineer to join our Infrastructure Engineering team and keep the healthcare platform running with high availability. You will design, build, and operate resilient systems on AWS, own end-to-end incident response, and drive disaster-recovery exercises with real RTO/RPO targets.

You will strengthen observability, CI/CD pipelines, and IaC while working hands-on in Kubernetes.

Qualifications

  • 4+ years in Site Reliability Engineering or Cloud Infrastructure.
  • Deep, current expertise in AWS, Kubernetes, and distributed systems.
  • Experience running disaster recovery or failover exercises.
  • A track record of driving incident response and postmortems yourself.
  • Solid grounding in CI/CD automation, GitOps, and infrastructure as code.
  • Appetite for staying close to the system rather than one step removed.
  • Bonus: healthcare IT, EHR data, or HIPAA/SOC 2-governed environments.
  • Bonus: experience with high-traffic, mission-critical SaaS or IoT platforms.

Responsibilities

  • Design, build, and operate resilient, self-healing infrastructure across our AWS-based platform.
  • Own incident response end-to-end: detection, mitigation, root-cause investigation, and postmortems.
  • Design and run disaster-recovery and failover exercises with real RTO/RPO targets.
  • Build out observability with SLIs, SLOs, and alerting that teams trust.
  • Build and maintain CI/CD pipelines and infrastructure as code (Terraform, GitOps).
  • Work hands-on in Kubernetes, beyond dashboards.
  • Partner day-to-day with our platform lead and on-call ownership.
  • Shape our security and compliance posture related to protected health data.

Skills

SRE/Cloud Infra
AWS
Kubernetes
Disaster Recovery
Incident Response
CI/CD & GitOps
IaC
Systems Thinking

Tools

AWS
Kubernetes
Terraform
GitOps

Job description

About Kontakt.io

Inside health systems, where every second can matter, operations are still spread across dozens of disconnected tools and platforms. Kontakt.io is changing that.

We combine proprietary hardware, AI-powered intelligence, and deep integrations with the technology health systems already have in place to build real-time understanding of what's happening across their operations. That intelligence becomes the execution layer care teams have been missing, helping them make smarter decisions and deliver better patient care.

Backed by Goldman Sachs and trusted by leading health systems including HCA Healthcare, Sutter Health, AdventHealth, Trinity Health, Northwell Health, Cleveland Clinic, and the U.S. Department of Veterans Affairs, we’ve more than doubled our revenue and are rapidly scaling with a clear path toward $100M in annual recurring revenue.

If you're excited to solve hard problems and help health systems deliver better care, we'd love to meet you!

About the role

We're looking for a Senior Site Reliability Engineer to join our Infrastructure Engineering team and get their hands directly into the systems that keep our healthcare platform running for hospitals and care teams who can't afford downtime. This is a builder's seat — you'll carry real operational weight and have direct influence over how our infrastructure evolves.

What you'll do
  • Personally design, build, and operate resilient, self-healing infrastructure across our AWS-based platform
  • Own incident response end-to-end: detection, mitigation, root-cause investigation, and postmortems that actually change how the system behaves next time
  • Design and run disaster-recovery and failover exercises with real RTO/RPO targets — you'll be the one who knows exactly what happens when things break
  • Build out observability that's genuinely tuned — SLIs, SLOs, and alerting people trust, not noise
  • Build and maintain CI/CD pipelines and infrastructure as code (Terraform, GitOps)
  • Work hands-on in Kubernetes, below the abstraction layer — you'll know the system, not just the dashboard
  • Partner day-to-day with our platform lead, sharing real production ownership and on-call
  • Shape our security and compliance posture (HIPAA, SOC 2 Type 2) as it relates to infrastructure handling protected health data
What you bring
  • 4+ years in Site Reliability Engineering or Cloud Infrastructure
  • Deep, current expertise in AWS, Kubernetes, and distributed systems, with the depth to go past the vocabulary
  • Real experience running disaster recovery or failover exercises
  • A track record of driving incident response and postmortems yourself
  • Solid grounding in CI/CD automation, GitOps, and infrastructure as code
  • An appetite for staying close to the system rather than one step removed from it
  • Bonus: healthcare IT, EHR data, or HIPAA/SOC 2-governed environments
  • Bonus: experience with high-traffic, mission-critical SaaS or IoT platforms
Logistics, Perks & Benefits
  • Built for collaboration - our team a hybrid schedule of 3 days/week minimum from our New York City office
  • Equity in a high-growth company scaling toward $400M+ ARR and backed by leading investors
  • Full health, dental, and vision coverage, a 401k, paid time off, paid parental leave and all the tools you need to do your best work
  • Autonomy to solve meaningful problems with work that ships quickly and makes a difference
Compensation

The expected salary range for this role is $200,000 – $240,000 for New York-based candidates. Actual compensation within this range will be determined based on relevant experience, skills, and qualifications. In exceptional cases, where a candidate’s experience or qualifications significantly exceed those anticipated for this role, we may consider the candidate for a more senior level. This role may also be eligible for equity and bonus compensation.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Site Reliability Engineer
Senior Site Reliability Engineer

Kontakt.io • New York (NY)

Hybrid
USD 200,000 - 240,000
Hybrid schedule
Equity
Health, dental, and vision
+3
Lead Site Reliability Engineer
Lead Site Reliability Engineer

Kontakt.io • New York (NY)

Hybrid
USD 200,000 - 250,000
Hybrid work 3 days/week in NYC office.
Equity in a high-growth company
Health, dental, vision insurance
+1
Senior Software Engineer
Senior Software Engineer

Kontakt.io • New York (NY)

Hybrid
USD 200,000 - 250,000
Equity
Health insurance
401k
+1
Site Reliability Engineer
Site Reliability Engineer

Fabric Labs, Inc. • New York (NY), Northern (KY)

Hybrid
USD 135,000 - 160,000
Medical, dental, vision
Unlimited PTO
401(k) plan
+2
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

Jobgether • United States

Remote
USD 150,000 - 200,000
Competitive salary
Comprehensive healthcare coverage
401(k) plan with company matching
+3
Senior DevOps Engineer
Senior DevOps Engineer

eSolutionsFirst • Palo Alto (CA), Northern (KY)

On-site
USD 170,000 - 220,000
Equity
Staff Site Reliability Engineer
Staff Site Reliability Engineer

Pivotal Health • Los Angeles (CA)

Hybrid
USD 230,000 - 260,000
Competitive compensation
Full health, dental, vision
401(k) plan
+2
Senior Software Engineer - SRE
Senior Software Engineer - SRE

Socure • City of Albany (NY)

On-site
USD 160,000 - 180,000
Site Reliability Engineer
Site Reliability Engineer

Tandem Inc. • Lehi (UT), Northern (KY)

On-site
USD 120,000 - 180,000
Competitive salary
Meaningful equity
401k & insurance
+1
Senior Site Reliability Engineer
Senior Site Reliability Engineer

Storm2 • Scottsdale (AZ)

Hybrid
USD 140,000 - 150,000
Competitive healthcare, dental, and vision coverage
401(k) with company match
Generous PTO and paid holidays
+1