Senior Data Engineer

United States Digital Space LLC

United States

On-site

USD 120,000 - 180,000

Full time

5 days ago
Be an early applicant
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Strava is seeking a Senior Data Engineer to join our Data Team and build reliable, scalable data pipelines and models for analytics and business use cases.

You will shape data governance, protect privacy, and evolve our data warehouse and lake with dbt, Airflow, Spark, and cloud technologies. A hybrid model requires on-site work in San Francisco three days per week.

Qualifications

  • 3–5+ years of professional experience in Data Engineering, Data Infrastructure, Software Engineering, or a related field.
  • Experience with SQL and data modeling for large-scale analytical systems.
  • Experience building and operating ETL/ELT data processing systems with dbt, Airflow, Spark.
  • Proven ability to develop reusable tooling to improve pipelines and testing.
  • Proficiency in Python, Scala, Java, or Go.
  • Familiarity with modern data warehouses and lakes (Snowflake, Databricks, BigQuery, Redshift, Iceberg).
  • Understanding of data governance, lineage, retention, and GDPR/privacy requirements.

Responsibilities

  • Design, build, and operate foundational data systems and assets across analytical, operational, and business use cases.
  • Develop scalable ingestion and transformation frameworks for data lake and data warehouse environments.
  • Create reusable data engineering tools and abstractions around dbt and related frameworks.
  • Model durable domain data entities with clear semantics and ownership.
  • Build workflows supporting governance, privacy, and lifecycle management per GDPR.
  • Improve reliability and observability of the data platform with tests, lineage, and monitoring.
  • Optimize processing and storage for performance, cost, and scalability across warehouse and lake.
  • Collaborate with cross-functional teams to set data architecture and engineering standards.

Skills

SQL & modeling
Pipelines ownership
Programming languages
Data governance
Cloud platforms
Data warehousing

Tools

dbt
Airflow
Spark
Snowflake
Databricks
BigQuery
Redshift
Iceberg
Delta Lake
Kafka
Flink
Kubernetes
Data catalogs

Job description

About StravaStrava is the app for active people. With over 200 million athletes in more than 185 countries, it’s more than tracking workouts—it’s where people make progress together, from new habits to new personal bests. No matter your sport or how you track it, the company’s got you covered. Find your crew, crush your goals, and make every effort count. Start your journey with the company today.

Our mission is simple: to motivate people to live their best active lives. We believe in the power of movement to connect and drive people forward.

About This RoleWe are looking for a Senior Data Engineer to join our Data Team and help build reliable, scalable data systems that support analytics, data science, and critical business use cases across the company.

In this role, you will design and operate data pipelines, build high-quality domain data models, improve our dbt and data transformation workflows, and help ensure data is accurate, well-governed, and easy to use. You will also contribute to areas such as data ingestion, data quality, privacy and GDPR workflows, and the ongoing evolution of our data warehouse and data lake.

We follow a flexible hybrid model that translates to more than half of your time on-site in our San Francisco office — three days per week.

What You’ll Do
  • Design, build, and operate foundational data systems and shared data assets that serve a broad range of analytical, operational, and business use cases across the company.
  • Build and evolve scalable data ingestion and transformation frameworks that move, process, clean, standardize, and organize data across our data lake and data warehouse.
  • Develop reusable data engineering tools and abstractions that improve how engineers build and operate data pipelines, including frameworks and capabilities around technologies such as dbt.
  • Design high-quality, durable domain data models — such as user, subscription, activity, or other core business domains — that provide consistent definitions and reusable foundations for teams across the company.
  • Build systems and workflows that support data governance, privacy, and regulatory requirements, including GDPR-related deletion, retention, access, and data lifecycle management.
  • Improve the reliability and observability of our data platform through automated testing, data quality checks, lineage, monitoring, alerting, and operational tooling.
  • Optimize large-scale data processing and storage for performance, maintainability, scalability, and cost across both warehouse and data lake environments.
  • Partner with data engineers, analytics engineers, software engineers, data scientists, security, privacy, and infrastructure teams to establish scalable data architecture and engineering standards.
You will be successful here by
  • Thinking beyond individual pipelines and designing reusable systems, abstractions, and data models that solve common problems across multiple teams and use cases.
  • Building well-defined domain data assets with clear semantics, ownership, lineage, and interfaces so downstream consumers can confidently build on top of them.
  • Applying strong data modeling principles to represent complex business entities and relationships in ways that are extensible, understandable, and efficient.
  • Maintaining a high bar for data correctness, reliability, privacy, and operational excellence across critical production data systems.
  • Making thoughtful engineering tradeoffs across data freshness, scalability, storage, compute cost, complexity, and developer productivity.
  • Proactively identifying recurring pain points in the data development lifecycle and creating tooling or platform capabilities that eliminate manual work and improve engineering velocity.
  • Designing data systems with governance and regulatory requirements in mind, rather than treating privacy and compliance as downstream concerns.
  • Bringing software engineering discipline to data infrastructure through testing, modular design, version control, CI/CD, observability, documentation, and code review.
What You’ll Bring to the Team
  • You have 3–5+ years of professional experience in Data Engineering, Data Infrastructure, Software Engineering, or a related field, with experience owning production data systems.
  • You have strong expertise in SQL and data modeling, including experience designing dimensional, normalized, or domain-oriented data models for large-scale analytical systems.
  • You have experience building and operating ETL/ELT and data processing systems using technologies such as dbt, Airflow, Spark, or similar frameworks.
  • You have experience developing reusable tooling, frameworks, or abstractions that improve how data pipelines and transformations are built, tested, deployed, or operated.
  • You are proficient in at least one general-purpose programming language such as Python, Scala, Java, or Go and are comfortable applying software engineering principles to data systems.
  • You understand modern data warehouse and data lake architectures and have worked with technologies such as Snowflake, Databricks, BigQuery, Redshift, Iceberg, Delta Lake, or similar systems.
  • You have experience processing and transforming large datasets, including handling schema evolution, data normalization, deduplication, backfills, incremental processing, and data quality.
  • You understand data governance and data lifecycle concepts such as lineage, retention, deletion, access control, PII handling, and GDPR/privacy requirements.
  • You have experience implementing production-grade data quality, monitoring, alerting, testing, and observability for data pipelines and datasets.
  • You can independently reason about data architecture and make sound technical decisions around modeling, ingestion, transformation, storage, reliability, scalability, and maintainability.
  • You are comfortable working with cloud infrastructure such as AWS, GCP, or Azure and understand the infrastructure that supports large-scale data processing systems.

Experience with Kafka, Flink or other streaming systems; Kubernetes; Iceberg or other open table formats; data catalogs and lineage systems; schema management; CDC; or internal developer platforms for data engineering is a plus.

Why Join Us?

Movement brings us together. At the company, we’re building the world’s largest community of active people, helping them stay motivated and achieve their goals.

Our global team is passionate about making movement fun, meaningful, and accessible to everyone. Whether you’re shaping the technology, growing our community, or driving innovation, your work at the company makes an impact.

When you join the company, you’re not just joining a company—you’re joining a movement. If you’re ready to bring your energy, ideas, and drive, let’s build something incredible together.

the company builds software that makes the best part of our athletes’ days even better. Just as we’re deeply committed to unlocking their potential, we’re dedicated to providing a world-class, inclusive workplace where our employees can grow and thrive, too. We’re backed by Sequoia Capital, TCV, Madrone Partners and Jackson Square Ventures, and we’re expanding in order to exceed the needs of our growing community of global athletes. Our culture reflects our community. We are continuously striving to hire and engage teammates from all backgrounds, experiences and perspectives because we know we are a stronger team together.

the company is an equal opportunity employer. In keeping with the values of the company, we make all employment decisions including hiring, evaluation, termination, promotional and training opportunities, without regard to race, religion, color, sex, age, national origin, ancestry, sexual orientation, physical handicap, mental disability, medical condition, disability, gender or identity or expression, pregnancy or pregnancy-related condition, marital status, height and/or weight.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

Socket.dev • San Francisco (CA)

Hybrid
USD 150,000 - 210,000
Senior Data Engineer
Senior Data Engineer

User Interviews • San Francisco (CA)

Hybrid
USD 140,000 - 210,000
Senior Data Engineer
Senior Data Engineer

Strava, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Hybrid work model in SF
Senior Data Engineer
Senior Data Engineer

Strava • San Francisco (CA)

Hybrid
USD 150,000 - 190,000
Senior Data Engineer
Senior Data Engineer

TOGETHXR • San Francisco (CA), Northern (KY)

Hybrid
USD 140,000 - 190,000
Senior Server Engineer, Data Products
Senior Server Engineer, Data Products

United States Digital Space LLC • United States

Hybrid
USD 150,000 - 210,000
Senior Server Engineer, Data Products
Senior Server Engineer, Data Products

Apply • San Francisco (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Senior Server Engineer, Data Products
Senior Server Engineer, Data Products

Strava • San Francisco (CA)

On-site
USD 180,000 - 240,000
Senior Server Engineer, Data Products
Senior Server Engineer, Data Products

TOGETHXR • San Francisco (CA), Northern (KY)

Hybrid
USD 180,000 - 240,000
Senior Server Engineer, Data Products
Senior Server Engineer, Data Products

Strava, Inc. • San Francisco (CA), Northern (KY)

Hybrid
USD 190,000 - 270,000