Senior PySpark Data Engineer — ETL & Big Data

VDart

Irving (TX)

On-site

USD 120,000 - 150,000

Full time

38 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

VDart is seeking a PySpark Consultant with minimum 6+ years of experience for a full-time role in Irving, TX. The role focuses on building and maintaining PySpark-based ETL pipelines, ensuring data quality, and optimizing distributed data workflows.

Strong coding skills in Python and SQL, plus experience with Hadoop, Hive, and Kafka, are essential to succeed. You will lead data processing efforts, document code lineage, and collaborate with data ingestion and analytics teams while adhering to

Qualifications

  • +4+ years of experience in big data development, Hadoop, Hive & Spark framework.
  • Strong Python, PySpark Development and SQL knowledge.
  • Certification in big data or cloud technologies is preferred.
  • Experience with SAS is a plus.

Responsibilities

  • Design, develop and optimize PySpark ETL pipelines.
  • Ensure data quality and integrity across data processing workflows.
  • Troubleshoot and resolve issues in PySpark applications and workflows.
  • Understand source data, dependencies and data flow from converted PySpark code.
  • Integrate PySpark code with frameworks such as Ingestion Framework, DataLens, etc.
  • Ensure compliance with data security, privacy regulations, and organizational standards.

Skills

PySpark
Python
SQL
Hadoop
Hive
Kafka
Data Warehousing
CI/CD
DevOps
Leadership
Communication

Tools

Hadoop
Hive
Kafka

Job description

VDart is seeking a PySpark Consultant with minimum 6+ years of experience for a full-time role in Irving, TX. The role focuses on building and maintaining PySpark-based ETL pipelines, ensuring data quality, and optimizing distributed data workflows.

Strong coding skills in Python and SQL, plus experience with Hadoop, Hive, and Kafka, are essential to succeed. You will lead data processing efforts, document code lineage, and collaborate with data ingestion and analytics teams while adhering to

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior PySpark Data Engineer
Senior PySpark Data Engineer

VDart Inc • Irving (TX)

On-site
USD 120,000 - 160,000
PySpark Consultant
PySpark Consultant

VDart • Irving (TX)

On-site
USD 120,000 - 150,000
Senior PySpark Data Engineer: ETL & Scalable Pipelines
Senior PySpark Data Engineer: ETL & Scalable Pipelines

Covetus • Irving (TX)

On-site
USD 100,000 - 130,000
Senior PySpark Data Engineer: Build Robust ETL Pipelines
Senior PySpark Data Engineer: Build Robust ETL Pipelines

MTK Technologies • Irving (TX)

On-site
USD 90,000 - 120,000
Senior Databricks Data Engineer – PySpark & ETL
Senior Databricks Data Engineer – PySpark & ETL

OpenTalent • United States

On-site
USD 120,000 - 160,000
Senior PySpark Data Engineer - Hadoop & ETL
Senior PySpark Data Engineer - Hadoop & ETL

Tata Consultancy Services • Charlotte (NC)

On-site
USD 100,000 - 110,000
Discretionary annual incentive
Comprehensive medical coverage
Family leaves
+4
Senior ETL SDET - PySpark Expert (Remote)
Senior ETL SDET - PySpark Expert (Remote)

Ethereum Technologies LLC • New York (NY), Northern (KY)

Hybrid
USD 83,000 - 124,000
Data Engineer
Data Engineer

The Value Maximizer • United States

On-site
USD 90,000 - 120,000
PySpark Consultant
PySpark Consultant

VDart Inc • Irving (TX)

On-site
USD 120,000 - 160,000
Spark Data Engineer (PySpark) - Cloud ETL & Pipelines
Spark Data Engineer (PySpark) - Cloud ETL & Pipelines

Tata Consultancy Services • Irving (TX)

On-site
USD 90,000 - 120,000
Discretionary Annual Incentive
Medical Coverage (Health, Dental & Vis
401K Plan
+1