Senior Data Platform Engineer

Crustdata (YC F24)

San Francisco (CA)

On-site

USD 140,000 - 200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

A growing tech company in San Francisco is looking for a Software Engineer to design and build scalable data systems. The ideal candidate has at least 3 years of experience in software engineering, particularly in data engineering. You will develop robust data pipelines and support data science initiatives, using modern tools and technologies. A collaborative, fast-paced startup mentality is essential for this role.

Qualifications

  • 3+ years of professional software engineering experience focused on data engineering.
  • Strong programming skills in Python or modern languages like Java or Go.
  • Experience with big data technologies and pipeline orchestration tools.

Responsibilities

  • Design, build, and maintain data infrastructure and analytics platforms.
  • Develop and scale data pipelines (ETL/ELT).
  • Implement workflow orchestration for daily data jobs.

Skills

Python
Big Data technologies (Spark, Flink, Dask)
Workflow management tools (Airflow, Dagster, Prefect)
Problem-solving
Collaboration

Tools

Kafka
Docker
Kubernetes

Job description

This range is provided by Crustdata (YC F24). Your actual pay will be based on your skills and experience — talk with your recruiter to learn more.

Base pay range

$140,000.00/yr - $200,000.00/yr

The Role

We are looking for a foundational member of our engineering team: a highly motivated Software Engineer to own the design, creation, and evolution of our data platform. You will be part of the team that owns the data ingestion and management infrastructure that powers Crustdata’s capabilities.

If you are passionate about building robust, scalable data systems and want to see your work directly influence customers, this is the role for you.

What You'll Do
  • Architect & Build: Design, build, and maintain our core data infrastructure, including our data warehouse and data lake, using modern cloud technologies (AWS, GCP, or Azure).
  • Pipeline Development: Develop and scale robust, fault-tolerant data pipelines (ETL/ELT) to ingest and process massive volumes of structured and unstructured data from diverse sources.
  • Enable Data Science & ML: Create the foundational platform to support our data scientists and ML engineers. This includes building systems for feature engineering, model training, and deploying ML models into production.
  • Orchestration at Scale: Implement and manage workflow orchestration for hundreds of daily data jobs, ensuring reliability, monitorability, and efficiency using tools like Airflow, Dagster, or Prefect.
  • Real-time Infrastructure: Build and manage real-time data streaming pipelines using technologies like Kafka or Flink to power live dashboards and time-sensitive product features.
  • Data Quality & Governance: Champion data quality and reliability. Implement frameworks for data validation, testing, and monitoring to ensure our data is accurate and trustworthy.
Who You Are
  • Experience: You have 3+ years of professional software engineering experience, with a significant focus on data engineering or building backend systems at scale.
  • Strong Coder: You possess strong programming skills in Python or another modern language (e.g., Java, Go).
  • Big Data Expertise: You have hands-on experience with modern big data technologies such as Spark, Flink, or Dask.
  • Pipeline Orchestration: You have practical experience with workflow management tools like Temporal, Airflow, Dagster, or Prefect.
  • Problem Solver: You are a pragmatic problem-solver who can navigate ambiguity, manage complexity, and take ownership of projects from inception to completion.
  • Startup Mentality: You are excited to work in a fast-paced, collaborative environment and wear multiple hats.
Nice to Haves
  • Experience with real-time streaming technologies (Kafka, Pulsar, Kinesis).
  • Familiarity with containerization and orchestration (Docker, Kubernetes).
  • Knowledge of modern data warehousing and lakehouse architectures (e.g., Delta Lake, Iceberg).
Seniority level

Mid-Senior level

Employment type

Full-time

Job function

Engineering and Information Technology

Industries: Software Development

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Platform Engineer
Senior Data Platform Engineer

United States Digital Space LLC • Boston (MA)

On-site
USD 150,000 - 215,000
Senior Data Engineer | Emerging Products
Senior Data Engineer | Emerging Products

United States Digital Space LLC • United States

Hybrid
USD 110,000 - 200,000
Medical insurance
Dental insurance
Vision insurance
+3
Senior Data Platform Engineer
Senior Data Platform Engineer

Selby Jennings • Oakland (CA)

On-site
USD 140,000 - 210,000
Sr Data Engineer
Sr Data Engineer

Instrumentl • United States

On-site
USD 120,000 - 160,000
Sr Data Engineer
Sr Data Engineer

INSPYR Solutions • Plano (TX)

On-site
USD 150,000 - 170,000
Senior Data Platform Engineer
Senior Data Platform Engineer

Harnham • Los Angeles (CA)

On-site
USD 120,000 - 150,000
Data Engineer
Data Engineer

UpRecruit • Los Angeles (CA)

Remote
USD 130,000 - 150,000
Software Engineer (Data Platform)
Software Engineer (Data Platform)

The Recruiting Guy • Boston (MA)

Remote
USD 125,000 - 200,000
Senior Data Engineer
Senior Data Engineer

Harnham • Richardson (TX)

Hybrid
USD 140,000 - 160,000
Hybrid schedule (4 days onsite)
Career growth in a stable enterprise
Modern cloud technologies & tools
+1
Senior Data Engineer
Senior Data Engineer

Uneek Global • Massachusetts

Hybrid
USD 115,000 - 170,000
Free home tech setup
Wellness benefits
Hybrid work flexibility