Backend Engineer, High-Volume Data Processing

AMBSIMS & Associates Corp

New York (NY)

On-site

USD 220,000 - 300,000

Full time

3 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Benefits offered by this job

Relocation assistance
Visa transfers supported

Job summary

AMBSIMS & Associates Corp seeks a backend data engineer to own production data pipelines that turn raw enterprise data into clean, de-identified datasets for AI labs. You’ll work in a small team with ownership and autonomy, building multi-stage pipelines in Python on AWS and ensuring correctness and reliability.

You’ll apply NLP/NER techniques for de-identification, develop internal dashboards to monitor pipeline health, and ship features quickly while maintaining high data quality.

Qualifications

  • A relevant undergraduate STEM degree from a top-30 U.S. institution.
  • 5-10 years of software engineering experience with ownership of production backend systems.
  • Hands-on experience building multi-stage, high-volume data-processing pipelines.
  • Strong fundamentals in distributed and asynchronous systems: queues, workers, concurrency, idempotency, retries, partial failure, and recovery.
  • Meaningful full-time startup experience with ambiguity and ownership (internships don’t count).
  • Strong skills in a general-purpose backend language; Python on AWS; Go, Rust, or Java welcome.
  • Experience with AWS.
  • Familiarity with NLP and NER.
  • Willingness to work onsite in Dumbo, Brooklyn, 5 days a week; relocation available.

Responsibilities

  • Design, build, and own production multi-stage data-processing pipelines end to end.
  • Build systems that stay correct under failure, with retries, partial failures, and recovery.
  • Apply NLP and Named Entity Recognition (NER) techniques to de-identify and transform enterprise data.
  • Build internal tooling and dashboards that show pipeline performance and failures.
  • Work across the stack, in AWS and mainly Python, to ship quickly and improve reliability and quality.
  • Make pragmatic technical decisions with minimal process and high autonomy.

Skills

Python
AWS
Go
Rust
Java
NLP/NER
Distributed systems
Data pipelines
Ownership

Education

Bachelor’s degree in STEM (top-30 U.S. institution)

Tools

AWS

Job description

Location: Dumbo, Brooklyn, NY (onsite, 5 days/week)

Type: Full-time

Compensation: $220K-$300K base salary

Relocation: Up to $10K

Visa: Open to visa transfers (including OPT and H-1B). Additional sponsorship may be considered for the right candidate.

Openings: 2

About Client

Client is the primary source of real, proprietary enterprise data for the world's leading frontier AI labs. They acquire enterprise data generated through collaboration, communication, and building. They transform it into de-identified datasets that remain useful, and license those datasets directly to frontier AI labs.

They've scaled from $0 to a multi-eight-figure run rate in a matter of months. They have about 14 people, with a lean engineering team of around five, backed by Floodgate, Afore Capital, Ludlow Ventures, and Hustle Fund. The client was built by the team behind Sunset, which scaled to an eight-figure run rate and helped hundreds of venture-backed startups wind down.

The Role

You’ll own the backend systems that turn raw enterprise data into clean, de-identified, high-value datasets. This is not a role about moving data between systems. The work is code applied to the data itself: multi-stage pipelines that process large volumes of messy, real-world data where correctness matters as much as throughput.

You’ll join a small team where each engineer has real ownership, the problems are ambiguous, and your work directly affects revenue.

What You’ll Do
  • Design, build, and own production multi-stage data-processing pipelines end to end
  • Build systems that stay correct under failure, with sound handling of retries, partial failures, and recovery
  • Apply NLP and Named Entity Recognition (NER) techniques to de-identify and transform enterprise data
  • Build internal tooling and dashboards that show how the pipeline is performing and where it’s breaking
  • Work across the stack, in AWS and mostly Python, to ship quickly and improve reliability and quality
  • Make pragmatic technical decisions with minimal process and a lot of autonomy

Required

  • A relevant undergraduate STEM degree from a top-30 U.S. institution
  • 5 -10 years of software engineering experience, with personal ownership of production backend systems
  • Hands-on experience building and running multi-stage, high-volume data-processing pipelines
  • Strong fundamentals in distributed and asynchronous systems: queues, workers, concurrency, idempotency, retries, partial failure, and recovery. We care about correctness and recovery under failure, not just scale.
  • Meaningful full-time experience at an early-stage or high-growth startup, with real ambiguity and ownership (internships and brief stints don’t count)
  • Strong skills in a general-purpose backend language. Our stack is Python on AWS, but strong Go, Rust, or Java engineers are welcome.
  • Experience with AWS
  • Familiarity with NLP and NER
  • Willingness to work onsite in Dumbo, Brooklyn, 5 days a week. We support relocation for candidates outside NYC.

Nice to have

  • Brand-name company experience paired with a meaningful startup stint
  • A clear record of career progression
  • Batch or workflow orchestration experience
  • A backend-leaning full-stack background, especially internal tools and dashboards
Compensation & Logistics
  • Base salary: $220K-$300K
  • Relocation support: up to $10K
  • Visa transfers supported, including OPT and H-1B
  • Onsite 5 days/week in Dumbo, Brooklyn
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Backend Engineer — High-Volume Data Processing
Backend Engineer — High-Volume Data Processing

Morgan Pinnacle Group • New York (NY)

On-site
USD 220,000 - 300,000
Relocation support
Visa sponsorship
Competitive equity
+1
Full Stack Engineer
Full Stack Engineer

jobr.pro • New York (NY)

On-site
USD 190,000 - 215,000
Full Stack Engineer
Full Stack Engineer

talentpluto • New York (NY)

On-site
USD 190,000 - 260,000
Senior Backend Software Engineer
Senior Backend Software Engineer

Ketch • San Francisco (CA)

On-site
USD 160,000 - 220,000
Full medical/dental/vision
401(k)
Equity
+1
Backend Engineer — Onsite, High-Volume Data Pipelines
Backend Engineer — Onsite, High-Volume Data Pipelines

Morgan Pinnacle Group • New York (NY)

On-site
USD 220,000 - 300,000
Relocation support
Visa sponsorship
Competitive equity
+1
Back End Developer
Back End Developer

Mentor Talent Acquisition • New York (NY)

On-site
USD 100,000 - 130,000
Back End Developer
Back End Developer

FORT • United States

On-site
USD 140,000 - 210,000
Meaningful equity
Remote across US & Canada
Software Engineer, Data Platform
Software Engineer, Data Platform

Rebar • New York (NY)

On-site
USD 140,000 - 180,000
agentic tooling budget
lunches provided, dinners provided (as
great culture and office banter
Backend Engineer — Data Pipeline
Backend Engineer — Data Pipeline

Sunset • Northern (KY), New York (NY)

Hybrid
USD 120,000 - 180,000
Senior Full Stack Engineer
Senior Full Stack Engineer

talentpluto • New York (NY)

On-site
USD 210,000 - 260,000