ML Data Engineer - Pipelines, Datasets & Quality

Sesame

San Francisco (CA)

On-site

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

401 (k) with 3.5% employer match
100% employer-paid health, vision, and dental benefits
Unlimited PTO and sick time
Flexible spending account with employer matching
Employee Assistance Program (EAP)
Competitive stock options

Job summary

Sesame is seeking a Data Engineer to build and maintain data pipelines crucial for AI models in San Francisco. You will collaborate with machine learning engineers to ensure access to structured data for model training and evaluation.

This role focuses on developing production pipelines for complex data including voice and conversational data, underlining strong SQL and Python skills. At Sesame, you will contribute significantly to data workflows and governance.

Qualifications

  • 5+ years in data engineering, with experience supporting ML or AI teams.
  • Strong SQL and Python skills.
  • Experience with modern data platforms and tooling.

Responsibilities

  • Design and build production data pipelines for model training.
  • Partner with ML engineers to understand data requirements.
  • Develop data quality frameworks to ensure data integrity.

Skills

SQL
Python
Data quality frameworks
ETL/ELT pipelines
Workflow orchestration systems
Communication skills

Tools

Airflow
Kubernetes
Ray
Spark

Job description

Sesame is seeking a Data Engineer to build and maintain data pipelines crucial for AI models in San Francisco. You will collaborate with machine learning engineers to ensure access to structured data for model training and evaluation.

This role focuses on developing production pipelines for complex data including voice and conversational data, underlining strong SQL and Python skills. At Sesame, you will contribute significantly to data workflows and governance.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

ML Data Engineer: Scalable Pipelines & Data Quality
ML Data Engineer: Scalable Pipelines & Data Quality

Sesame • San Francisco (CA)

On-site
USD 120,000 - 160,000
401(k) max employer match: 3.5% of compensation
100% employer-paid health, vision, and dental benefits
Unlimited PTO and sick time
+3
ML Data Engineer - DataOps & Pipelines
ML Data Engineer - DataOps & Pipelines

X Development, LLC • Mountain View (CA)

On-site
USD 166,000 - 244,000
Bonus
Equity
Benefits
Data Engineer, Machine Learning
Data Engineer, Machine Learning

Sesame • San Francisco (CA)

On-site
USD 120,000 - 160,000
401(k) max employer match: 3.5% of compensation
100% employer-paid health, vision, and dental benefits
Unlimited PTO and sick time
+3
ML Data Engineer: Preference Data Pipelines
ML Data Engineer: Preference Data Pipelines

vizcom • United States

On-site
USD 140,000 - 190,000
Data Engineer: Scalable Pipelines & Analytics Models
Data Engineer: Scalable Pipelines & Analytics Models

Lazer Logistics • Alpharetta (GA)

On-site
USD 105,000 - 130,000
General Benefits
AI Data Engineer: ML Pipelines, LLMs & MLOps
AI Data Engineer: ML Pipelines, LLMs & MLOps

SolomonEdwards • Charlotte (NC)

On-site
USD 120,000 - 180,000
ML Data Engineer: Preference Data Pipelines
ML Data Engineer: Preference Data Pipelines

Doist • San Francisco (CA)

On-site
USD 250,000 - 400,000
Lead Multimodal Data Engineer for ML Pipelines
Lead Multimodal Data Engineer for ML Pipelines

Hark, Inc. • San Jose (CA)

On-site
USD 170,000 - 450,000
Data Engineer, Machine Learning
Data Engineer, Machine Learning

Sesame • San Francisco (CA)

On-site
USD 120,000 - 160,000
401 (k) with 3.5% employer match
100% employer-paid health, vision, and dental benefits
Unlimited PTO and sick time
+3
Data & AI/ML Engineer: Build Scalable ML Pipelines
Data & AI/ML Engineer: Build Scalable ML Pipelines

System • New York (NY)

Hybrid
USD 100,000 - 140,000