Senior Data Platform SRE — Scale & Observability

Datavant

Tallahassee (FL)

On-site

USD 168,000 - 200,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Datavant is seeking a Senior Site Reliability Engineer to join the Data & ML Platform team. You will own the Databricks and Snowflake platforms lifecycle, build scalable, observable infrastructure, and drive CI/CD for data pipelines and ML workflows.

You will collaborate with Data Scientists, Analysts, and App Engineers to deliver secure, self-service, production-grade solutions while optimizing costs and ensuring reliability in a hybrid cloud environment.

Qualifications

  • 6+ years in SRE, platform engineering, or DevOps roles.
  • Hands-on Databricks experience with workspace setup, cluster/job mgmt, and CI/CD integration.

Responsibilities

  • Operate and improve Databricks and Snowflake platforms including automation and cost optimization.
  • Design for reliability across cloud environments with failover and capacity planning.
  • Advance observability with monitoring, alerting, and logging; define SLOs/SLAs.

Skills

SRE experience
Cloud infrastructure
Observability
CI/CD

Tools

Databricks
Snowflake
AWS
Datadog
GitHub Actions
Terraform

Job description

Datavant is seeking a Senior Site Reliability Engineer to join the Data & ML Platform team. You will own the Databricks and Snowflake platforms lifecycle, build scalable, observable infrastructure, and drive CI/CD for data pipelines and ML workflows.

You will collaborate with Data Scientists, Analysts, and App Engineers to deliver secure, self-service, production-grade solutions while optimizing costs and ensuring reliability in a hybrid cloud environment.

Get your free, confidential resume review.

or drag and drop your file here.