Senior Data Engineer

ActivTrak

Austin, Northern (TX, KY)

Hybrid

USD 130,000 - 170,000

Full time

4 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

ActivTrak is expanding its data engineering team to design and operate production data pipelines at petabyte scale. You’ll own the platform, focusing on orchestration, deployment, observability, and reliability to enable ML and analytics work.

This is an individual contributor role with broad ownership, remote within the United States, and limited travel. Strong Python, SQL, and ETL/ELT design are essential.

Qualifications

  • 5+ years designing and operating data pipelines in production at scale.
  • Strong Python and production-grade software engineering practices.
  • Demonstrably strong SQL and data modeling skills.
  • Experience with batch and streaming data processing and ETL/ELT design.

Responsibilities

  • Own the production data pipelines and orchestration for a scalable data platform.
  • Design, implement, and operate data ingestion, cleansing, and transformation.
  • Build reliable deployment, observability, and monitoring for pipelines.
  • Collaborate across teams while retaining individual contributor ownership.

Skills

Python
SQL
Data Pipelines
ETL/ELT

Tools

Spark
Docker
Kubernetes

Job description

We're growing our data engineering team to design and operate the production systems that make our machine learning and analytics work scalable, observable, reliable, and economical. This is a platform-ownership role: you're building the infrastructure a petabyte-scale data platform runs on, not executing tickets handed down from another team.

You'll work on problems like:
  • Designing and operating production data pipelines across a layered (bronze/silver/gold) data platform: sourcing, cleansing, and transforming data at scale
  • Solving incremental sync and stream processing as data volume and account count grow — a real, unsolved scaling problem for us today
  • Building regional data architecture that respects data residency and PII boundaries (e.g., data belonging to a region can't simply be centralized into one global store)
  • Designing the production systems, orchestration, and deployment mechanisms that turn Data Science's models and analysis into durable, monitored, reliable pipelines
Where this role starts and ends:

Data Science owns problem formulation, features, model and scoring logic, evaluation, and model performance. You own the pipeline infrastructure, orchestration, deployment mechanisms, observability, scalability, and operational reliability that put their work into production. The primary ownership is clear, but you'll work together across that boundary when production issues span model and platform.

This is an individual contributor role with substantial ownership of the platform. It does not include people-management responsibilities.

Must-Haves:
  • 5+ years designing and operating data pipelines in production at scale
  • Experience with batch and stream data processing, including incremental/streaming architectures
  • Strong Python and production-grade software engineering practices (testing, code review, version control, monitoring)
  • Demonstrably strong SQL
  • Experience with ETL/ELT pipeline design and orchestration
Nice-to-Haves:
  • Experience with regional/multi-region data storage and data residency constraints
  • Parallel dataframes (Dask, Spark, or similar)
  • API design/implementation (microservices, REST, etc.)
  • Cloud environment experience (GCP or AWS), Docker/Containers, Kubernetes
  • Machine learning deployment or serving experience
Why Should You Apply?
  • Own the platform problem that determines how fast and how reliably ActivTrak's ML and analytics can scale
  • Meaningful scaling challenges: incremental/streaming processing, regional data residency, and petabyte-scale data across hundreds of millions of events per day
  • Small, senior team with real ownership and visibility to leadership
Work environment
  • Position is remote within US
  • Minimal travel
  • Limited physical demands

This is an incredible opportunity to embark on an exciting journey with a dynamic, VC-backed company. If you have a proven track record of creative thinking, a drive for learning, and a deep commitment to collaboration, we want to talk to you!

ActivTrak is an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. ActivTrak does not discriminate in employment on the basis of race, color, religion, sex, national origin, political affiliation, sexual orientation, marital status, disability, or age.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Scientist
Senior Data Scientist

ActivTrak • Austin (TX), Northern (KY)

Hybrid
USD 130,000 - 170,000
Senior Data Platform Engineer — Production Pipelines (Remote)
Senior Data Platform Engineer — Production Pipelines (Remote)

ActivTrak • Austin (TX), Northern (KY)

Hybrid
USD 130,000 - 170,000
Sr Data Engineer
Sr Data Engineer

Instrumentl • United States

On-site
USD 120,000 - 160,000
Senior Data Engineer
Senior Data Engineer

ICE • Atlanta (GA)

On-site
USD 100,000 - 130,000
Senior Data Platform Engineer
Senior Data Platform Engineer

Harnham • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Senior Data Platform Engineer
Senior Data Platform Engineer

Selby Jennings • Oakland (CA)

On-site
USD 140,000 - 210,000
Senior Data Engineer
Senior Data Engineer

The Trade Desk • Ventura (CA)

On-site
USD 125,000 - 229,000
Senior Data Engineer | Emerging Products
Senior Data Engineer | Emerging Products

United States Digital Space LLC • United States

Hybrid
USD 110,000 - 200,000
Medical insurance
Dental insurance
Vision insurance
+3
Senior Data Engineer
Senior Data Engineer

Assembl • New York (NY)

On-site
USD 140,000 - 190,000
Lead Data Platform Engineer
Lead Data Platform Engineer

Tykhe Inc • New York (NY)

On-site
USD 180,000 - 240,000