Data Engineer

Salma Health, Inc.

San Francisco (CA)

Hybrid

USD 119,000 - 185,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, vision benefits
Discretionary bonuses
Paid time off (PTO)

Job summary

Salma Health, Inc. is seeking a Data Engineer to build the data backbone for mental and behavioral health practice. This hands-on role requires experience in data pipelines, with a focus on AWS and dbt.

Candidates with 4-7 years of experience who can maintain orchestration layers and have solid Python and SQL skills will be preferred. The position offers remote flexibility, a base salary between $119,000 and $185,000, and a variety of benefits.

Qualifications

  • 4-7 years of professional experience building and operating data pipelines in production.
  • Strong Python skills for module writing, structuring code, and debugging.
  • Solid SQL skills for performance reasoning and window functions.

Responsibilities

  • Build the data pipeline from third-party APIs to metrics for clinical teams.
  • Maintain and improve orchestration layer with Dagster.
  • Write tests and contribute to documentation.

Skills

Python
SQL
Data pipelines
dbt
AWS
Git

Tools

Dagster
GraphQL
CloudFormation

Job description

Data Engineer
Location

Remote. Hybrid role. Preference for candidates located in the San Francisco Bay Area, San Diego, or Salt Lake City. Remote is possible, with the expectation of regular in‑person collaboration.

Employment Type

Full time

Department

Technology

We are looking to hire a Data Engineer to join our team as we build the data backbone for a mental and behavioral health practice. This role will build the platform that turns appointments, assessments, billing, and patient engagement data into the metrics our clinical and operations teams rely on. As a mid‑level data engineer, you’ll own meaningful pieces of our pipeline end‑to‑end: from pulling data out of third‑party APIs, through medallion architecture transformations in dbt, to exposing curated metrics through our semantic layer.

This is a hands‑on role on a small team. You’ll write code that runs in production every day, ship improvements weekly, and have direct visibility into how the data is used. We work in a HIPAA‑regulated environment, so thoughtfulness about data handling is part of the job.

What You’ll Work On
  • Maintaining and improving the orchestration layer: Dagster assets, jobs, schedules, sensors, and the dependency graph that ties extraction → loading → transformation together.
  • Adding new data sources to the pipeline; extracting from APIs (GraphQL, REST), Google Drive folders, and CSV/JSONL drops on S3, then landing them in our bronze schemas via Dagster assets.
  • Building silver and gold dbt models that transform raw source data into our unified entity model following the medallion architecture.
  • Extending our semantic layer so business metrics are available to downstream consumers (BI tool dashboards, AI agents, ad‑hoc analysis) without re‑deriving logic.
  • Operating the platform on AWS: ECS Fargate services, RDS, S3, Secrets Manager, CloudFormation templates, and the CodePipeline‑based CI/CD that deploys our data platform. All of our data platforms are deployed with IaC tools.
  • Writing tests (pytest for Python, dbt tests for models, data quality tests) and contributing to internal documentation as new patterns emerge.
What We’re Looking For
Required
  • 4‑7 years of professional experience building and operating data pipelines in production.
  • From conversation to shipped data product: you’re comfortable owning a request end‑to‑end: scoping it with a non‑technical stakeholder, writing requirements clear enough that you (and others) can build against them, implementing the models or metrics, and verifying with the stakeholder that what shipped solves their problem.
  • Strong Python: comfortable writing modules, structuring code for reuse and testability, and debugging issues across an async or orchestrated pipeline.
  • Solid SQL skills, including window functions, CTEs (including recursive ones), and the ability to reason about query performance.
  • Hands‑on experience with dbt: building models, writing tests, and understanding materializations.
  • Working knowledge of an orchestration framework: (Dagster, Airflow, Prefect, or similar), including the mental model of assets/tasks, dependencies, and scheduling.
  • Comfort with AWS fundamentals: S3, IAM, Secrets Manager, and either ECS or Lambda for compute.
  • Git‑based workflows: code review, and writing PRs that are reviewable.
Nice to Have
  • Experience with Dagster specifically.
  • Experience with semantic layer tools (Cube.js, dbt Semantic Layer/MetricFlow, LookML).
  • Healthcare data experience (HIPAA, EHR systems, ICD‑10/CPT codes).
  • CloudFormation, Terraform, or another IaC tool.
  • Experience with GraphQL APIs as a consumer (pagination, introspection, dealing with rate limits and retries).
  • Familiarity with identity resolution patterns or slowly‑changing dimension modeling.
How We Work
  • Small, focused team; your work ships and gets used quickly.
  • Pragmatic engineering: we favor readable code, clear naming conventions, and well‑documented patterns over clever abstractions. Our internal "how to add X" guides are first‑class artifacts.
  • Tests on everything; CI runs dbt parse, dg check defs, and pytest on every PR.
Compensation & Benefits
  • Base: The base salary range for this role is $119,000–$185,000, depending on geographic location, experience, and qualifications. Salma Health uses a tiered compensation structure based on candidate location. Specific range details are available during the interview process.
  • Incentives: Discretionary bonus based on company and individual performance.
  • Benefits: Medical, dental, vision, PTO, and additional benefits.
Work Authorization

Sponsorship for employment authorization may be considered on a case‑by‑case basis depending on the role and candidate qualifications.

Equal Opportunity & Accessibility Statement

We are committed to providing a workplace that is inclusive, respectful, and free from discrimination. We welcome applicants of all backgrounds and make employment decisions without regard to race, color, religion, sex (including pregnancy, childbirth, and related medical conditions), sexual orientation, gender identity or expression, national origin, ancestry, citizenship, age, physical or mental disability, medical condition, genetic information, marital status, military or veteran status, or any other characteristic protected by California or federal law.

In accordance with the California Fair Chance Act, we will consider qualified applicants with arrest and conviction records.

If you require a reasonable accommodation during the application or hiring process, please contact us directly — we’re happy to help.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Salma Health • United States

Hybrid
USD 119,000 - 185,000
Medical insurance
Dental insurance
Vision insurance
+2
Data Engineer
Data Engineer

contexture • Arizona

Hybrid
USD 110,000 - 160,000
Data Engineer
Data Engineer

Courier Health • New York (NY)

On-site
USD 120,000 - 150,000
100% paid health benefits
401(k) with employer match
Unlimited Vacation
+3
Data Engineer, Senior (Hybrid)
Data Engineer, Senior (Hybrid)

Releady • San Francisco (CA)

Hybrid
USD 82,656 - 96,432
Data Engineer
Data Engineer

AndHealth • Columbus (OH)

On-site
USD 100,000 - 130,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
Software Engineer (Platforms & Integrations)
Software Engineer (Platforms & Integrations)

Salma Health • Sacramento (KY)

Hybrid
USD 123,000 - 183,000
Stock options
Discretionary bonus
Medical, dental, vision, PTO
Data Engineer
Data Engineer

AndHealth LLC • Columbus (OH)

On-site
USD 90,000 - 110,000
Medical Insurance
Dental Insurance
Vision Insurance
+2
Senior Data Engineer
Senior Data Engineer

Khealthcareers • New York (NY)

Hybrid
USD 150,000 - 200,000
Hybrid work schedule
18 vacation days
Stock options
+4
Senior Data Engineer
Senior Data Engineer

Adecco • Mesa (AZ)

On-site
USD 121,000 - 135,000
Analytics Engineer, Data Platform
Analytics Engineer, Data Platform

AndHealth • Columbus (OH)

On-site
USD 80,000 - 110,000
Medical, Dental, Vision Insurance
Paid time off
Short- and Long-Term Disability