Site Reliability Engineer (SRE), Data Products

United States Digital Space LLC

United States

Hybrid

USD 135,000 - 160,000

Full time

41 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Health benefits
Generous PTO
401(k) matching
Parental leave
Hybrid work schedules
Career development

Job summary

United States Digital Space LLC is seeking a passionate Site Reliability Engineer to own reliable deployment and operation of cloud infrastructure across multiple AWS accounts and Snowflake connectivity, supporting our autism and IDD care software platform.

Join a fast-moving startup with hybrid offices in Fort Lauderdale, FL; Holmdel, NJ; and Verona, Italy, offering competitive compensation, comprehensive benefits, generous PTO, 401(k) matching, parental leave, and career development.

Qualifications

  • 3–5 years of SRE/DevOps experience with AWS.
  • Strong Linux admin, networking, and incident troubleshooting.
  • Experience with automation, CI/CD, and IaC is preferred.

Responsibilities

  • Support deployment, configuration, and daily operations across AWS accounts and Linux workloads.
  • Administer and troubleshoot S3, Glue, Kinesis, PrivateLink, and security groups.
  • Investigate incidents, outages, and connectivity issues to restore service quickly.
  • Automate deployment and runbooks; maintain reproducible processes.
  • Collaborate with engineering, security, and data teams and document procedures.

Skills

SRE/DevOps
Linux administration
AWS (S3, Glue, Kinesis, PrivateLink)
Cloud networking
Snowflake connectivity
Bash/Python
CI/CD & IaC
Monitoring/incident response

Tools

AWS
Snowflake
CI/CD
Terraform/IaC

Job description

the company is a leading provider of autism and IDD care software for Applied Behavior Analysis (ABA), multidisciplinary therapy, and special education. Trusted by more than 200,000 users, we enable therapy providers, educators, and employers to scale the way they deliver ABA and related therapies with innovative technology, market-leading industry expertise, and world-class customer satisfaction.

We are looking for a passionate Site Reliability Engineer (SRE) to join our team! You will be a trusted partner responsible for reliable deployment and daily operation of cloud infrastructure and data-platform connectivity across a growing set of autism and IDD care software solutions.

As a key member of the Engineering team, you will take ownership of supporting several AWS accounts and their Linux-based workloads and managed services, including Amazon S3, AWS Glue, Amazon Kinesis, AWS PrivateLink, and security groups. You will also support Snowflake account connectivity and deployment processes, working closely with developers, data engineers, security, and other technical and non-technical partners to keep environments secure, available, and ready for change.

If you have a passion for cloud infrastructure, reliability, automation, and technology in general, enjoy and thrive in an agile, fast-moving, ever-changing startup environment, welcome operational and technical challenges of all shapes and sizes, have excellent interpersonal skills and a sense of humor, and enjoy rolling up your sleeves and jumping in, then read on!

Key Accountabilities:
  • Support deployment, configuration, and daily operations across several AWS accounts and the Linux servers and workloads running in those environments.
  • Administer, monitor, and troubleshoot Amazon S3 buckets, AWS Glue jobs, Amazon Kinesis streams, AWS PrivateLink connectivity, security groups, and related AWS networking and access dependencies.
  • Investigate and resolve infrastructure incidents, deployment failures, service interruptions, and connectivity issues, with a focus on restoring service quickly and preventing recurrence.
  • Maintain and improve repeatable deployment and operational processes for AWS infrastructure and workloads through automation, version‑controlled configuration, and clear runbooks.
  • Support secure connectivity to Snowflake accounts, including troubleshooting network paths, private connectivity, endpoint configuration, and environment‑specific access issues.
  • Support and improve Snowflake deployment processes across environments so that configuration and platform changes are consistent, reviewable, and reliable.
  • Collaborate closely with application engineering, data engineering, security, and other teams; contribute to daily stand‑up meetings and document operational procedures, known issues, and recovery steps.
Desired Skills and Experience:
  • 3-5 years of experience in Site Reliability Engineering, DevOps, Cloud Engineering, or production infrastructure operations, with hands‑on responsibility for AWS environments.
  • Strong Linux administration and troubleshooting skills, including system services, networking, permissions, logs, and command‑line diagnostics.
  • Hands‑on working knowledge of AWS services used by the role, including Amazon S3, AWS Glue, Amazon Kinesis, AWS PrivateLink, security groups, and multi‑account operational practices.
  • Good understanding of cloud networking fundamentals, including DNS, TCP/IP, routing, private endpoints, security groups, TLS, and troubleshooting application‑to‑service connectivity.
  • Experience supporting Snowflake connectivity and deployment processes, including diagnosing access and network issues and coordinating changes across multiple environments.
  • Experience with scripting and operational automation using tools such as Bash or Python; familiarity with CI/CD and Infrastructure as Code practices is highly desirable.
  • Experience with monitoring, logging, alerting, incident troubleshooting, and creating practical operational documentation and runbooks.
  • Well‑organized and self‑motivated, with strong multi‑tasking skills and a positive, professional approach to working across technical and non‑technical teams.
Base Salary Range

$135,000—$160,000 USD

Backed by Roper Technologies, Inc. (Nasdaq: ROP), the company is entering an exciting phase of growth, innovation, and scale.

Recognized as one of the best places to work over 10 times by organizations such as Inc, Built In, and NJBIZ, our culture is centered around impact, inclusion, and flexibility.

As a hybrid company with collaborative offices in Ft. Lauderdale, FL; Holmdel, NJ; and Verona, Italy, we foster a workplace where top talent can thrive and make a real difference in the lives of those we serve.

  • competitive compensation
  • comprehensive health benefits
  • generous PTO
  • 401(k) matching
  • paid parental leave to our full‑time employees.
  • hybrid work schedules
  • career development support
  • wellness programs
  • opportunities to give back through CR Cares™
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Site Reliability Engineer (SRE), Data Products
Site Reliability Engineer (SRE), Data Products

Centralreach-8 • Holmdel Township (NJ)

On-site
USD 135,000 - 160,000
Hybrid work model
Comprehensive health benefits
Generous PTO
+3
Site Reliability Engineer (SRE), Data Products New Holmdel, New Jersey
Site Reliability Engineer (SRE), Data Products New Holmdel, New Jersey

CentralReach, LLC • Holmdel Township (NJ)

On-site
USD 135,000 - 160,000
Competitive compensation
Health benefits
Generous PTO
+6
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

United States Digital Space LLC • United States

Hybrid
USD 160,000 - 180,000
Hybrid work model
401(k) matching
Parental leave
+1
Principal Platform Engineer
Principal Platform Engineer

United States Digital Space LLC • United States

Remote
USD 200,000 - 215,000
Health benefits by company
Generous PTO
401(k) matching
+4
SRE – Data Platform & Cloud Reliability
SRE – Data Platform & Cloud Reliability

United States Digital Space LLC • United States

Hybrid
USD 135,000 - 160,000
Health benefits
Generous PTO
401(k) matching
+3
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

CentralReach • Holmdel Township (NJ)

Hybrid
USD 160,000 - 180,000
Health benefits
PTO
401(k) matching
+2
Sr. Site Reliability Engineer
Sr. Site Reliability Engineer

CentralReach • Fort Lauderdale (FL)

Hybrid
USD 160,000 - 180,000
Hybrid work model
Health benefits
PTO & 401(k) matching
+1
Sr. Site Reliability Engineer New Holmdel, New Jersey
Sr. Site Reliability Engineer New Holmdel, New Jersey

CentralReach, LLC • Holmdel Township (NJ), Northern (KY)

On-site
USD 160,000 - 180,000
Health benefits
PTO and holidays
401(k) matching
+2
Technical Product Analyst
Technical Product Analyst

United States Digital Space LLC • United States

On-site
USD 85,000 - 105,000
Comprehensive health benefits
Generous PTO
401(k) matching
+3
Sr. Software Engineer, Ruby on Rails
Sr. Software Engineer, Ruby on Rails

CentralReach • United States

On-site
USD 120,000 - 180,000
competitive compensation
comprehensive health benefits
generous PTO
+6