Data Engineer, Alternative Data | Delta One Trading | Experienced Hire

Susquehanna International Group, LLP

New York (NY)

On-site

USD 225,000 - 250,000

Full time

3 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Susquehanna International Group, LLP is seeking a Data Engineer for a Systematic Delta One desk. You will partner with researchers and traders to transform raw vendor and alternative data into timely, research-ready datasets powering alpha research and backtests.

The role emphasizes Python data pipelines, scalable ingestion over S3/SFTP, and evolving datasets with high data quality. Collaboration and fast iteration are essential.

Qualifications

  • 5+ years of Python data applications and pipelines over large historical datasets.
  • Experience ingesting and normalizing third-party or vendor data at scale (S3/SFTP, data shares, APIs).
  • Strong SQL and familiarity with Parquet, Arrow, DuckDB, NumPy, Pandas, or Polars.
  • Understanding of data modeling, data accuracy, and reproducible research workflows; bi-temporal modeling preferred.
  • Experience operating production data pipelines: monitoring, backfill and restatement handling, incident forensics.
  • Experience applying LLMs to data problems or building LLM-assisted tooling is a plus.
  • Experience with cloud data delivery (AWS S3, Snowflake) is a plus; C++ is a plus.
  • Experience in quantitative finance or electronic trading environments is a plus; advanced degree a plus.

Responsibilities

  • Own end-to-end onboarding of new datasets from evaluation to production monitoring.
  • Build point-in-time-correct datasets preserving history and restatements.
  • Design entity-mapping datasets linking vendor identifiers to tradable instruments.
  • Extend platform to extract structure from unstructured vendor material and automate onboarding.
  • Run data-quality and vendor-evaluation studies informing trials and licensing.
  • Create research-ready datasets for large-scale historical analysis and backtesting.
  • Improve ingestion platform to onboard new datasets faster.

Skills

Python data pipelines
SQL / columnar tooling
Data modeling & reproducible research
Production data pipelines
LLMs / AI tooling
Cloud data delivery
Batch data processing (S3/SFTP)

Education

Advanced degree in CS/Math/Physics/Computer Engineering or related field

Tools

Parquet
Arrow
DuckDB
NumPy/Pandas/Polars

Job description

Overview

We are seeking a Data Engineer to join our Systematic Delta One desk, where engineers, researchers, and traders work side-by-side to develop scalable, fully automated trading strategies across liquid global products and venues. Partnering closely with our quantitative researchers, this role owns the path from raw vendor and alternative datasets to the point-in-time-correct, research-ready data that powers alpha research and signal development.

The ideal candidate combines strong Python engineering skills with hands‑on experience ingesting and normalizing third‑party data at scale: batch feeds over S3 and SFTP, cloud data shares, and APIs, and increasingly semi‑structured and unstructured sources such as documents, transcripts, and text. You will design pipelines and research tools that process billions of rows of historical data efficiently, reproducibly, and with a high degree of correctness.

A core part of the role is translating evolving research ideas into usable datasets and research infrastructure, working with our market‑data and compliance teams during vendor trials. Success in this role requires strong communication skills, intellectual curiosity, and the ability to iterate quickly as hypotheses and data requirements evolve.

How You'll Make an Impact:

  • Own the end-to-end onboarding of new vendor and alternative datasets: from evaluating samples and data dictionaries with researchers, through building ingestion pipelines, to production monitoring
  • Build point-in-time-correct datasets: preserving as‑delivered history and handling vendor restatements, revisions, and backfills, so backtests see exactly what was knowable at the time
  • Design entity‑mapping and reference datasets that connect vendor identifiers (brands, merchants, estimate line items) to tradable instruments
  • Extend the platform beyond tabular feeds: apply LLMs and agentic tooling to extract structure from unstructured vendor material (documents, filings, transcripts, data dictionaries) and to automate onboarding, entity‑resolution, and data‑quality workflows
  • Run data‑quality and vendor‑evaluation studies (coverage, revision behavior, panel stability) that directly inform trial and licensing decisions
  • Create research‑ready datasets optimized for large‑scale historical analysis and backtesting workflows
  • Improve the shared ingestion platform and tooling so that each new dataset onboards faster than the last
What we’re looking for
  • 5+ years of experience building Python data applications and pipelines over large historical datasets, with a performance‑aware mindset
  • Experience ingesting and normalizing third‑party or vendor data at scale (batch feeds over S3/SFTP, cloud data shares such as Snowflake, or APIs)
  • Strong SQL and familiarity with modern columnar and analytical tooling (Parquet, Arrow, DuckDB or similar), alongside NumPy, Pandas, or Polars
  • Strong understanding of data modeling, data accuracy, and reproducible research workflows; experience with temporal or versioned data (point‑in‑time, slowly changing dimensions, bitemporal modeling) strongly preferred
  • Demonstrated success operating production data pipelines: monitoring, alerting, backfill and restatement handling, incident forensics
  • Ability to work closely with researchers and scientists, taking ambiguous ideas and evolving them into robust datasets and scalable workflows
  • Experience applying LLMs to data problems (extraction, classification, entity resolution, data‑quality checking) or building LLM‑assisted and agentic tooling is a plus
  • Experience with cloud data delivery (AWS S3, Snowflake) is a plus; prior experience in C++ is a plus
  • Experience in quantitative finance or electronic trading environments is a plus but not required
  • An advanced degree in Computer Science, Mathematics, Physics, Computer Engineering, or a related field is a plus

What you can expect from us:

Real Impact: You will onboard the datasets that decide which signals get built, and see your pipelines feed research and production trading directly. Your work makes the whole research organization smarter, faster, and better.

Collaboration: Our data engineers, researchers, and traders work together daily; the feedback loop from a dataset you built to a strategy in production is short and visible.

Growth: We're looking for people who are naturally curious, relentless problem solvers, and have the desire to continuously innovate, learn, and grow; prior proprietary‑trading experience is not required.

About Susquehanna

Susquehanna is a global quantitative trading firm powered by scientific rigor, curiosity, and innovation. Our culture is intellectually driven and highly collaborative, bringing together researchers, engineers, and traders to design and deploy impactful strategies in our systematic trading environment. To meet the unique challenges of global markets, Susquehanna applies machine learning and advanced quantitative research to vast datasets in order to uncover actionable insights and build effective strategies. By uniting deep market expertise with cutting‑edge technology, we excel in solving complex problems and pushing boundaries together.

What we do

We are experts in trading essentially all listed financial products and asset classes, with a focus on derivatives trading. Through market making and market taking, we handle millions of trading transactions around the world every day, providing liquidity and ensuring competitive prices for buyers and sellers. While our presence in the market is broad, our trading desks are highly specialized, allowing for a deep understanding of unique drivers of each asset class.

The annual base pay range for this role is $225,000 - $250,000 + discretionary bonus + benefits. Susquehanna considers factors such as scope and responsibilities of the position, work experience, education/training, key skills, as well as market and organizational considerations when extending an offer.

#LI-DT1

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer, Alternative Data | Delta One Trading | Experienced Hire
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire

Socket.dev • New York (NY)

On-site
USD 225,000 - 250,000
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire

SIG Susquehanna • Northern (KY)

On-site
USD 140,000 - 200,000
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire

Susquehanna International Group • Bala Cynwyd (PA)

On-site
USD 110,000 - 150,000
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire
Data Engineer, Alternative Data | Delta One Trading | Experienced Hire

Susquehanna International Group, LLP • Lower Merion Township

On-site
USD 120,000 - 190,000
Data Engineer, Alternative Data for Systematic Trading
Data Engineer, Alternative Data for Systematic Trading

Susquehanna International Group, LLP • New York (NY)

On-site
USD 225,000 - 250,000
Data Engineer | Sports | Experienced Hire
Data Engineer | Sports | Experienced Hire

SIG Susquehanna • Pennsylvania

On-site
USD 150,000 - 210,000
Excellent salary and benefits
Opportunities to attend world sporting
Data Engineer | Sports | Experienced Hire
Data Engineer | Sports | Experienced Hire

SIG Susquehanna • Northern (KY)

Hybrid
USD 120,000 - 180,000
Data Engineer | Predictions | Experienced Hire
Data Engineer | Predictions | Experienced Hire

Susquehanna International Group, LLP • Lower Merion Township

On-site
USD 150,000 - 230,000
Excellent salary and benefits
Opportunities to attend world sporting
Reference Data Developer
Reference Data Developer

Quant Blueprint LLC • Sydney Township (ND)

On-site
USD 70,000 - 90,000
C++ Developer | Trading Strategies | Experienced Hire
C++ Developer | Trading Strategies | Experienced Hire

SIG Susquehanna • Northern (KY)

Hybrid
USD 140,000 - 190,000