Data Engineer, Analytics Data Products

The New York Times

New York (NY)

Hybrid

USD 110,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Medical, dental, and vision benefits
Flexible Spending Accounts (F.S.A.s)
401(k) plan with company match
Paid vacation
Paid parental leave

Job summary

A leading media organization in New York is seeking a Data Engineer to design and implement ELT/ETL pipelines. The role involves developing data transformations using dbt and PySpark, managing data storage on GCP and AWS, and ensuring data quality across systems. Ideal candidates will have 2+ years of experience in data engineering, proficiency in SQL and Python, and familiarity with cloud environments. This position offers a competitive salary range of $110,000 - $130,000 with flexible working arrangements.

Qualifications

  • 2+ years of hands-on experience in Data Engineering, Data Warehousing or equivalent.
  • Experience with production-level data modeling and Cloud Data Warehouses.
  • Familiarity with cloud services in GCP or AWS.

Responsibilities

  • Design and implement complex ELT/ETL pipelines.
  • Develop data transformations using dbt and PySpark.
  • Administer and tune Spark resources to optimize performance.

Skills

SQL
Python
dbt (data build tool)
PySpark
workflow orchestration tools (e.g., Airflow)
data modeling (dimensional modeling, Kimball)

Tools

BigQuery
GCP
AWS
Terraform

Job description

The mission of The New York Times is to seek the truth and help people understand the world. That means independent journalism is at the heart of all we do as a company. It’s why we have a world‑renowned newsroom that sends journalists to report on the ground from nearly 160 countries. It’s why we focus deeply on how our readers will experience our journalism, from print to audio to a world‑class digital and app destination. And it’s why our business strategy centers on making journalism so good that it’s worth paying for.

Responsibilities
  • Design, model, and implement complex ELT/ETL pipelines for the cleansed and curated data layers in the medallion architecture, taking full ownership of the data product’s structure, partitioning, documentation, and performance characteristics.
  • Develop advanced data transformations using dbt (data build tool) for relational data modeling and PySpark for large‑scale data processing within the Lakehouse, ensuring outputs meet strict Service Level Agreements and quality standards.
  • Collaborate across teams to define requirements and translate them into robust and scalable data models suitable for analytic consumption.
  • Manage the physical data storage across both GCP and AWS, selecting optimal file formats and designing efficient partitioning and clustering strategies.
  • Administer and tune Spark compute resources (e.g., Dataproc, EMR, or managed services) to optimize job execution time and cost.
  • Own core components of our centralized analytics environment, specifically focused on Hex, integrations, and the methods of data exposure and access controls; and support data activation strategies, ensuring seamless data consumption by analytic tools.
  • Optimize user queries and access patterns to maintain platform performance and cost efficiency.
  • Implement centralized data quality checks and observability mechanisms within the data pipeline to proactively identify and resolve data issues.
  • Contribute to the implementation of metadata management, data lineage, and role‑based access control (RBAC) initiatives across the Lakehouse environment.
  • Demonstrate support and understanding of our value of journalistic independence and a strong commitment to our mission to seek the truth and help people understand the world.
Basic Qualifications
  • 2+ years of hands‑on experience in a Data Engineering, Data Warehousing, Analytics Engineering or equivalent role.
  • Proficiency in SQL and experience with complex, production‑level data modeling (dimensional modeling, Kimball, OBT, or Data Vault).
  • Demonstrated experience designing, developing, and deploying end‑to‑end data products through the full Software Development Lifecycle.
  • Experience with a Cloud Data Warehouse, like BigQuery.
  • Proficiency in Python for scripting and data manipulation, including knowledge of PySpark or other Spark APIs.
  • Familiarity with cloud services and data storage components in at least one major cloud provider (GCP or AWS).
  • Experience with workflow orchestration tools (e.g., Airflow, Cloud Composer, or Prefect) and version control systems (Git).
Preferred Qualifications
  • Experience operating in a dual‑cloud environment (GCP/AWS).
  • Experience with Infrastructure‑as‑Code (IaC) tools like Terraform.
  • Experience with advanced Lakehouse file formats like Iceberg or Delta Lake.
  • Familiarity with experimentation or A/B testing platforms and the data required to support them.
  • Experience in data product quality standards through integration advanced testing, quality checks, and monitoring into the CI/CD pipeline.
Compensation & Benefits

The annual base pay range for this role is between $110,000 – $130,000 USD. For roles in the U.S., dependent on your role, you may be eligible for variable pay, such as an annual bonus and restricted stock. Benefits may include medical, dental and vision benefits, Flexible Spending Accounts (F.S.A.s), a company‑matching 401(k) plan, paid vacation, paid sick days, paid parental leave, tuition reimbursement and professional development programs. For roles outside of the U.S., information on benefits will be provided during the interview process.

Equal Opportunity Employment

The New York Times Company is committed to being the world’s best source of independent, reliable and quality journalism. To do so, we embrace a diverse workforce that has a broad range of backgrounds and experiences across our ranks, at all levels of the organization. We encourage people from all backgrounds to apply. We are an Equal Opportunity Employer and do not discriminate on the basis of an individual's sex, age, race, color, creed, national origin, alienage, religion, marital status, pregnancy, sexual orientation or affectional preference, gender identity and expression, disability, genetic trait or predisposition, carrier status, citizenship, veteran or military status and other personal characteristics protected by law. All applications will receive consideration for employment without regard to legally protected characteristics. The U.S. Equal Employment Opportunity Commission (EEOC)’s Know Your Rights Poster is available here . The New York Times Company will provide reasonable accommodations as required by applicable federal, state, and/or local laws. Individuals seeking an accommodation for the application or interview process should email reasonable.accommodations@nytimes.com. Emails sent for unrelated issues, such as following up on an application, will not receive a response.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer, Analytics Data Products
Data Engineer, Analytics Data Products

Thenewyorktimes • New York (NY)

Hybrid
USD 110,000 - 130,000
Medical, dental, and vision benefits
401(k) plan with company matching
Paid vacation and sick days
Data Engineer, Analytics Data Products
Data Engineer, Analytics Data Products

The New York Times • New York (NY)

Hybrid
USD 110,000 - 130,000
Medical, dental, and vision benefits
401(k) plan with company matching
Paid parental leave
+2
Data Engineer, Analytics Data Products New York, NY
Data Engineer, Analytics Data Products New York, NY

New York Times • California (MO)

On-site
USD 110,000 - 130,000
Annual bonus
Company-matching 401(k) plan
Paid parental leave
+1
Analytics Data Platform Engineer
Analytics Data Platform Engineer

New York Times • California (MO)

Hybrid
USD 110,000 - 130,000
Senior Analytics Engineer, Data Platform
Senior Analytics Engineer, Data Platform

Thenewyorktimes • New York (NY)

On-site
USD 124,000 - 135,000
Medical, dental, and vision benefits
401(k) plan
Paid vacation and sick days
+2
Senior Data Engineer, Customer-Facing Data Products
Senior Data Engineer, Customer-Facing Data Products

The New York Times • New York (NY)

Hybrid
USD 140,000 - 155,000
Medical, dental, and vision benefits
401(k) plan
Paid vacation and sick days
+2
Senior Analyst, Data & Insights, Growth (Algorithmic Targeting)
Senior Analyst, Data & Insights, Growth (Algorithmic Targeting)

Thenewyorktimes • New York (NY)

On-site
USD 101,000 - 114,000
Annual bonus
401(k) matching
Tuition reimbursement
+2
Senior Analyst, Data & Insights
Senior Analyst, Data & Insights

New York Times • New York (NY)

On-site
USD 101,000 - 114,000
Medical insurance
Dental insurance
Vision insurance
+8
Senior Data Engineer, Customer-Facing Data Products
Senior Data Engineer, Customer-Facing Data Products

Thenewyorktimes • New York (NY)

Hybrid
USD 140,000 - 160,000
Medical, dental, and vision coverage
Flexible Spending Accounts
401(k) plan
+2
Technical Product Manager II, Data Platforms New York, NY
Technical Product Manager II, Data Platforms New York, NY

New York Times • New York (NY)

On-site
USD 120,000 - 142,000