Full-Stack Engineer Senior

Donnelly-

New York (NY)

On-site

USD 140,000 - 200,000

Full time

5 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Donnelly- in New York seeks an experienced Senior Data Engineer to design scalable data platforms powering analytics and AI-driven solutions.

You will implement Lakehouse patterns using Databricks, Delta Lake, and Snowflake, and build streaming pipelines with Kafka and AWS. Collaboration with data scientists and stakeholders is essential for enterprise-grade data products.

Qualifications

  • Bachelor's or Master's degree in Computer Science, Engineering, or related field.
  • 8+ years of software or data engineering experience in enterprise-scale environments.
  • Proven ability to design and operate scalable data pipelines and platforms.

Responsibilities

  • Design, develop, and optimize scalable batch and real-time data pipelines.
  • Build data ingestion, transformation, and processing for diverse datasets.
  • Implement Lakehouse architectures with Databricks, Delta Lake, and Bronze-Silver-Gold patterns.
  • Develop streaming pipelines with Kafka, Spark Structured Streaming, and AWS Kinesis.
  • Collaborate with architects, data scientists, and stakeholders to deliver enterprise data solutions.
  • Ensure data quality, governance, and security across ecosystems.
  • Leverage Generative AI and LLMs to enhance data tooling and productivity.

Skills

Python
PySpark
Spark SQL
SQL
Kafka
Databricks
Delta Lake
Snowflake
Hive
AWS
Java

Education

Bachelor's or Master's degree in CS/Engineering

Tools

Airflow
Git
CI/CD
Spark
AWS Console
Tableau
Docker

Job description

We are seeking an experienced Senior Data Engineer to design, develop, and optimize scalable data platforms that power enterprise analytics, machine learning, and AI-driven solutions. The ideal candidate will bring deep expertise in modern data engineering practices, cloud-native architectures, Lakehouse platforms, and distributed data processing technologies.

This role will play a critical part in building reliable, high-performance data ecosystems leveraging Databricks, Spark, Delta Lake, Snowflake, Kafka, and AWS, while also contributing to the adoption of Generative AI, Large Language Models (LLMs), Retrieval Augmented Generation (RAG), and AI-assisted engineering solutions.

Key Responsibilities
Data Platform Engineering
  • Design, build, and maintain scalable batch and real-time data pipelines.
  • Develop and optimize data ingestion, transformation, and processing frameworks for structured, semi-structured, and unstructured datasets.
  • Implement modern Lakehouse architectures utilizing Databricks, Delta Lake, and Medallion (Bronze, Silver, Gold) design patterns.
  • Build data solutions that support enterprise analytics, reporting, and machine learning initiatives.
  • Ensure data quality, governance, lineage, security, and compliance across data ecosystems.
Big Data & Streaming Solutions
  • Develop distributed data processing applications using PySpark and Spark SQL.
  • Build and maintain streaming pipelines using Kafka, Spark Structured Streaming, and AWS Kinesis.
  • Design fault-tolerant, scalable systems capable of processing large data volumes with low latency.
  • Optimize workload performance through partitioning strategies, clustering, caching, and query tuning.
Cloud & Lakehouse Architecture
  • Architect and implement cloud-based data solutions on AWS.
  • Utilize AWS services including S3, EMR, EC2, Athena, Redshift, RDS, Lambda, IAM, SNS, and SQS.
  • Design data storage and processing strategies that maximize reliability while minimizing operational costs.
  • Support migration initiatives from traditional Hadoop and EMR environments to modern cloud-native platforms.
Data Operations & Automation
  • Develop orchestration and scheduling frameworks using Airflow and Databricks Workflows.
  • Build CI/CD pipelines and automation frameworks for deployment, monitoring, and data platform operations.
  • Collaborate closely with architects, analysts, data scientists, and business stakeholders to deliver enterprise-grade solutions.
AI & Intelligent Platform Engineering
  • Implement Generative AI-powered solutions for engineering productivity and operational excellence.
  • Develop applications leveraging Large Language Models (LLMs), Retrieval Augmented Generation (RAG), Vector Databases, and Model Context Protocol (MCP).
  • Build AI-assisted documentation, developer productivity tooling, and intelligent platform capabilities.
  • Evaluate emerging AI technologies and identify opportunities for adoption within data engineering processes.
Required Qualifications
  • Bachelor's or Master's degree in Computer Engineering, Computer Science, Information Systems, or a related field.
  • 8+ years of experience in software engineering, data engineering, or big data platform development.
  • Strong experience designing and implementing enterprise-scale data pipelines.
  • Hands-on expertise with:
    • Python
    • PySpark
    • Spark SQL
    • SQL
    • Kafka
    • Databricks
    • Delta Lake
    • Snowflake
    • Hive
  • Experience building data solutions on AWS cloud platforms.
  • Strong understanding of distributed computing, data modeling, and large-scale data processing.
  • Experience with Git-based development workflows and CI/CD practices.
  • Excellent analytical, troubleshooting, and problem-solving skills.
Preferred Qualifications
  • Experience with real-time streaming architectures and event-driven systems.
  • Knowledge of data governance, metadata management, and data quality frameworks.
  • Experience with generative AI technologies including:
    • LLMs
    • RAG
    • Vector Databases
    • AI Agents
    • MCP integrations
  • Experience developing developer productivity tools and AI-assisted engineering workflows.
  • Exposure to enterprise supply chain, retail, healthcare, or manufacturing data domains.
  • AWS certifications are highly preferred.
Technical Skills
Programming Languages
  • Python
  • Java
  • SQL
  • Shell Scripting
  • C/C++
Big Data & Data Engineering
  • PySpark
  • Spark SQL
  • Hive
  • Databricks
  • Delta Lake
  • Snowflake
  • Kafka
  • HBase
  • Sqoop
Workflow & Orchestration
  • Apache Airflow
  • Databricks Workflows
  • Oozie
Cloud Technologies
  • AWS S3
  • EMR
  • EC2
  • Athena
  • Redshift
  • RDS
  • IAM
  • Lambda
  • SNS
  • SQS
AI & Modern Engineering
  • Generative AI
  • Large Language Models (LLMs)
  • Retrieval Augmented Generation (RAG)
  • Agentic AI Systems
  • Model Context Protocol (MCP)
  • Vector Databases
Visualization & Tools
  • Tableau
  • Git
  • Docker
  • Splunk
  • IntelliJ IDEA
  • PyCharm
  • Cursor
Preferred Certifications
  • AWS Certified Solutions Architect - Associate
  • AWS Certified Cloud Practitioner
  • Databricks Certifications (preferred)
What Success Looks Like
  • Deliver highly scalable and reliable data pipelines.
  • Improve platform performance, efficiency, and cost optimization.
  • Enable enterprise-wide analytics and AI initiatives through trusted data products.
  • Drive modernization of data platforms and adoption of cloud-native architectures.
  • Leverage AI technologies to enhance engineering efficiency, automation, and innovation.
Ideal Candidate Profile:

A senior-level data engineer with extensive experience in Databricks, Spark, AWS, Kafka, Snowflake, and Lakehouse architectures, who is equally passionate about modern AI technologies and building intelligent data platforms for the future.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Sr. Data Engineer, Data Platform
Sr. Data Engineer, Data Platform

Mirion Technologies • United States

On-site
USD 120,000 - 160,000
Senior Data Engineer
Senior Data Engineer

Peyton Resource Group • Houston (TX)

On-site
USD 120,000 - 150,000
Senior Data Engineer
Senior Data Engineer

Tata Consultancy Services • Irvine (CA)

On-site
USD 120,000 - 180,000
Databricks Developer
Databricks Developer

Tredence Inc. • San Jose (CA)

On-site
USD 140,000 - 190,000
Senior Data Engineer
Senior Data Engineer

Prosum • Glendale (CA)

On-site
USD 150,000 - 210,000
Senior Data Engineer
Senior Data Engineer

Motion Recruitment Partners, LLC • Harrisonburg (VA)

On-site
USD 110,000 - 160,000
Bonus eligible
Medical, Dental, and Vision Insurance
Vacation Time
Senior Data Engineer-5
Senior Data Engineer-5

REALIGN LLC • Irvine (CA)

On-site
USD 140,000 - 210,000
Databricks Data Engineer
Databricks Data Engineer

Compunnel, Inc. • Spring (TX)

On-site
USD 110,000 - 140,000
Data Engineer
Data Engineer

Prodigy Resources • Denver (CO)

On-site
USD 110,000 - 170,000
Sr. Data Engineer
Sr. Data Engineer

Veritas Search Group • Glendale (CA)

Hybrid
USD 140,000 - 190,000