Data Infrastructure Engineer - Large-Scale ML Pipelines

Visa Hunt

San Francisco (CA)

On-site

USD 350,000 - 475,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Health, dental, and vision benefits
Unlimited PTO
Parental leave
Relocation support

Job summary

Thinking Machines in San Francisco is seeking an engineer to join our data infrastructure team. You will help architect and scale core infrastructure behind distributed training pipelines, multimodal data catalogs, and processing systems handling petabytes of data.

You'll collaborate with researchers to accelerate experiments, build high-throughput data ingestion and quality checks, and ensure traceability and reliability across the data lifecycle.

Qualifications

  • Bachelor’s degree or equivalent in computer science, engineering, or similar.
  • Proficiency in at least one backend language (Python or Rust).
  • Experience with distributed compute frameworks (Spark or Ray).
  • Familiarity with cloud infrastructure, data lake architectures, and batch and streaming pipelines.
  • Ability to own projects end-to-end.
  • Strong collaboration across cross-functional teams.
  • Bias for action and initiative to work across stacks to ship features.

Responsibilities

  • Design, build, and operate scalable, fault-tolerant infrastructure for LLM Research: distributed compute, data orchestration, and storage across modalities.
  • Develop high-throughput systems for data ingestion, processing, and transformation — including training data catalogs, deduplication, quality checks, and search.
  • Build systems for traceability, reproducibility, and robust quality control at every stage of the data lifecycle.
  • Implement and maintain monitoring and alerting to support platform reliability and performance.
  • Collaborate with research teams to unlock new features, improve data quality, and accelerate training cycles.

Skills

Python
Rust
Distributed compute
Data pipelines
Cloud infrastructure
End-to-end ownership
Collaborative teamwork

Education

Bachelor’s degree in CS or engineering

Tools

Spark
Ray
Kafka
dbt
Terraform
Airflow

Job description

Thinking Machines in San Francisco is seeking an engineer to join our data infrastructure team. You will help architect and scale core infrastructure behind distributed training pipelines, multimodal data catalogs, and processing systems handling petabytes of data.

You'll collaborate with researchers to accelerate experiments, build high-throughput data ingestion and quality checks, and ensure traceability and reliability across the data lifecycle.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Infrastructure Engineer — Scalable ML Data Pipelines | Flexible Hours
Data Infrastructure Engineer — Scalable ML Data Pipelines | Flexible Hours

JobCubby • San Francisco (CA), Northern (KY)

Hybrid
USD 500,000 - 850,000
Visa sponsorship
Office in San Francisco
Flexible working hours
Data Infrastructure Engineer - Scalable ML Pipelines
Data Infrastructure Engineer - Scalable ML Pipelines

Alljoined • San Francisco (CA)

On-site
USD 140,000 - 180,000
Competitive equity compensation
Options for housing support
Visa sponsorship
+2
Data Infrastructure Engineer for ML Pipelines
Data Infrastructure Engineer for ML Pipelines

Luma • Redwood City (CA)

On-site
USD 150,000 - 240,000
Staff Data Infrastructure Engineer for Scalable ML Pipelines
Staff Data Infrastructure Engineer for Scalable ML Pipelines

Neura Market • San Francisco (CA), Northern (KY)

Hybrid
USD 500,000 - 850,000
Competitive compensation
Equity donation matching
Generous vacation
+3
Senior ML Data Infrastructure Engineer - Scalable Pipelines
Senior ML Data Infrastructure Engineer - Scalable Pipelines

Jobtailor • California (MO)

On-site
USD 120,000 - 160,000
Data Engineering Lead: Scalable ML Data Pipelines
Data Engineering Lead: Scalable ML Data Pipelines

Hark • San Jose (CA)

On-site
USD 170,000 - 450,000
Data Foundations Engineer – Scalable Pipelines & Open AI
Data Foundations Engineer – Scalable Pipelines & Open AI

Reflection • San Francisco (CA)

On-site
USD 150,000 - 210,000
Top-tier compensation
Stock options
Health & wellness
+3
Staff Data Infrastructure Engineer: Scale AI Pipelines
Staff Data Infrastructure Engineer: Scale AI Pipelines

Inception • San Francisco (CA)

On-site
USD 140,000 - 190,000
Staff Infra Engineer, ML Data Pipelines & GPUs
Staff Infra Engineer, ML Data Pipelines & GPUs

Sieve • San Francisco (CA)

On-site
USD 150,000 - 210,000
401k
Health insurance
Meals & snacks
+2
Member of Technical Staff, Data Infrastructure
Member of Technical Staff, Data Infrastructure

Inception • San Francisco (CA)

On-site
USD 140,000 - 190,000