Forward Deployed Engineer, Data as a Service

Snorkel AI

New York (NY)

On-site

USD 150,000 - 320,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Equity options
Benefits

Job summary

Snorkel AI is seeking an engineering delivery professional to own the data pipeline lifecycle from data generation to ML-assisted workflows. You will drive end-to-end delivery, collaborating with the DaaS Delivery Operations team and cross-functional partners to design evaluation workflows, and implement quality standards that scale.

This role emphasizes technical leadership, turning ambiguous customer needs into production-grade assets, and delivering measurable impact across projects and

Qualifications

  • 3+ years in data science, data engineering, ML engineering or similar roles.
  • Strong Python and SQL data tooling experience.
  • Experience with ML/LLM in production and validation workflows.
  • Ability to translate ambiguous requirements into technical specs and workflows.
  • Experience integrating systems via APIs and building custom tooling.

Responsibilities

  • Build and deploy evaluators; design quality measurement systems.
  • Generate synthetic datasets to accelerate client engagements.
  • Package and deliver production-grade datasets with documentation and QA.
  • Configure and build custom applications for non-standard client requirements.
  • Define production specifications and workflows; enable go-live transitions.
  • Provide ongoing technical support to Delivery Managers and resolve blockers.

Skills

Python
SQL
ML engineering
API integration
Data pipelines

Tools

Pandas

Job description

About Snorkel

At Snorkel, we believe meaningful AI doesn’t start with the model, it starts with the data. We’re on a mission to help enterprises transform expert knowledge into specialized AI at scale. The AI landscape has gone through incredible changes since 2015, when Snorkel started as a research project in the Stanford AI Lab, to the generative AI breakthroughs of today. But one thing has remained constant: the data you use to build AI is the key to achieving differentiation, high performance, and production-ready systems. We work with some of the world’s largest organizations to empower scientists, engineers, financial experts, product creators, journalists, and more to build custom AI with their data faster than ever before.

About the Role - Multiple Levels

This is an engineering delivery role where you’ll translate ambiguous customer needs into production‑grade data, evaluation, and ML‑assisted workflows. This is a high‑impact role focused on end‑to‑end ownership of the AI data pipeline lifecycle, including developing and deploying ML‑based workflows and building the technical foundations that make our human‑in‑the‑loop (HITL) data generation and review faster, more reliable, and more effective.

You’ll work at the critical intersection of data science, data engineering, AI engineering and operations, partnering closely with our DaaS Delivery Operations team and cross‑functional stakeholders. You’ll develop technical specifications, design evaluation workflows, implement quality standards, measurement frameworks, and ML‑assisted applications which improve our data pipelines and unblock projects through technical innovation.

This role is ideal for someone who is comfortable working throughout the delivery lifecycle, rolling up their sleeves to solve complex multi‑faceted problems, thrives as a technical communicator and works well as a key member of a team.

Main Responsibilities
Project Execution & Delivery
  • Build and deploy evaluators, design and implement quality measurement systems to validate project outputs and ensure deliverables meet client expectations
  • Generate synthetic datasets by developing or adapting existing pipelines to accelerate client engagements and augment training data
  • Package and deliver production‑grade datasets with standardized formatting, comprehensive documentation, and quality assurance
  • Configure and build custom applications and off‑platform solutions for non‑standard or specialized client requirements
Production & Technical Partnership
  • Define production specifications and workflows, securing technical alignment with client teams to enable seamless go‑live transitions
  • Provide ongoing technical support to Delivery Managers, addressing complex questions, resolving technical blockers, and supporting customer rebuttals
  • Maintain specification consistency and alignment across customer and internal teams throughout the engagement lifecycle
  • Identify and document workflow best practices and automation opportunities, collaborating with DaaS Engineering to continuously improve delivery capabilities
Technical Leadership & Innovation
  • Maintain solution leaderboards and execute custom model benchmarking on existing datasets to demonstrate technical capabilities
  • Drive continuous improvement of technical assets, evaluation frameworks, and delivery processes to enhance speed, quality, and scalability
What We’re Looking For
  • 3+ years of experience in data science, data engineering, ML engineering, solutions engineering, forward‑deployed engineering, or other technical solution development roles.
  • Strong practical experience with Python and SQL data tooling required.
  • Familiarity with ML and LLM‑based solutions, including applying ML techniques in production contexts and building validation, evaluation, or quality measurement workflows for ML/LLM‑based systems.
  • Experience translating ambiguous customer or stakeholder requirements into technical specifications, workflows, and delivered solutions.
  • Experience integrating systems via APIs and building custom tooling or workflows for non‑standard technical requirements.
Compensation

The base salary range for this position is $150,000 – $320,000. This range reflects multiple factors, including role level and work location. The final offer within this range will be determined based on job‑related skills, experience, relevant education or training, interview performance, and other business considerations.

All offers also include equity in the form of employee stock options, as well as benefits.

Actual compensation will be determined based on factors including skills, qualifications, experience, and geographic location.

Salary range(s) for this role: $150,000 — $320,000 USD

Snorkel AI is proud to be an Equal Employment Opportunity employer and is committed to building a team that represents a variety of backgrounds, perspectives, and skills. Snorkel AI embraces diversity and provides equal employment opportunities to all employees and applicants for employment. Snorkel AI prohibits discrimination and harassment of any type on the basis of race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local law. All employment is decided on the basis of qualifications, performance, merit, and business need. We will ensure that individuals with disabilities are provided reasonable accommodation to participate in the job application or interview process, to perform essential job functions, and to receive other benefits and privileges of employment. Please contact us to request accommodation.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engagement Manager - Data as a Service
Engagement Manager - Data as a Service

Snorkel AI • Redwood City (CA), San Francisco (CA), New York (NY)

Hybrid
USD 140,000 - 230,000
Equity in the form of employee stock options
Comprehensive compensation package
Senior/Staff FDE - Synthetic Data Generation New New York City, NY (Hybrid); San Francisco, CA (Hybrid)
Senior/Staff FDE - Synthetic Data Generation New New York City, NY (Hybrid); San Francisco, CA (Hybrid)

Snorkel AI, Inc. • New York (NY), Northern (KY)

Hybrid
USD 180,000 - 320,000
Director, Forward Deployed Researcher - Data as a Service
Director, Forward Deployed Researcher - Data as a Service

Snorkel AI • New York (NY)

On-site
USD 228,000 - 404,000
Senior Software Engineer - Expert Data Collection Platform
Senior Software Engineer - Expert Data Collection Platform

Snorkel AI • San Francisco (CA), Redwood City (CA)

Hybrid
USD 192,000 - 240,000
Senior/Staff FDE - Synthetic Data Generation
Senior/Staff FDE - Synthetic Data Generation

Snorkel AI • New York (NY), San Francisco (CA)

Hybrid
USD 180,000 - 320,000
Equity offered
Benefits package
Director, Research - Evaluation & Training
Director, Research - Evaluation & Training

Snorkel AI • United States

On-site
USD 275,000 - 425,000
Director, Research - Evaluation & Training
Director, Research - Evaluation & Training

Snorkel AI • San Francisco (CA)

On-site
USD 275,000 - 425,000
Senior Manager, Forward Deployed Research
Senior Manager, Forward Deployed Research

Socket.dev • New York (NY)

Hybrid
USD 185,000 - 322,000
Lead Data Scientist
Lead Data Scientist

Snorkel AI • San Francisco (CA)

Hybrid
USD 130,000 - 200,000
Streamlit
Snowflake Cortex
Director, Research - Evaluation & Training
Director, Research - Evaluation & Training

Snorkel AI • New York (NY)

On-site
USD 275,000 - 425,000