Senior PySpark Data Engineer

VDart Inc

Irving (TX)

On-site

USD 120,000 - 160,000

Full time

12 days ago
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

VDart Inc. in Irving, TX is seeking an experienced Big Data Developer to design and implement PySpark-based data processing pipelines. You will ensure data quality, troubleshoot PySpark workflows, and integrate with ingestion and data-lens frameworks while maintaining data security standards.

You should have 4+ years in big data with Hadoop, Hive and Spark, strong Python/SQL skills, and a proven ability to lead, communicate, and document data lineage across complex data flows.

Qualifications

  • 4+ years of experience in big data development, Hadoop, Hive & Spark.
  • Strong Python and SQL knowledge.
  • Experience with PySpark and Spark frameworks.
  • Knowledge of data security and CI/CD practices.

Responsibilities

  • Design and implement ETL pipelines and data transformations.
  • Troubleshoot PySpark applications and workflows.
  • Ensure data quality and integrity across processing workflows.
  • Document data lineage and data flow.
  • Understand sources and dependencies of converted PySpark code.
  • Integrate PySpark code with Ingestion Framework and DataLens where applicable.

Skills

Big data development
Python
SQL
PySpark
Spark framework
Hadoop
Hive
Kafka
Data lineage
CI/CD / DevOps
Leadership
Communication

Tools

PySpark
Spark
Hadoop
Hive
Kafka

Job description

VDart Inc. in Irving, TX is seeking an experienced Big Data Developer to design and implement PySpark-based data processing pipelines. You will ensure data quality, troubleshoot PySpark workflows, and integrate with ingestion and data-lens frameworks while maintaining data security standards.

You should have 4+ years in big data with Hadoop, Hive and Spark, strong Python/SQL skills, and a proven ability to lead, communicate, and document data lineage across complex data flows.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior PySpark Data Engineer — ETL & Big Data
Senior PySpark Data Engineer — ETL & Big Data

VDart • Irving (TX)

On-site
USD 120,000 - 150,000
PySpark Consultant
PySpark Consultant

VDart • Irving (TX)

On-site
USD 120,000 - 150,000
Senior PySpark Data Engineer: ETL & Scalable Pipelines
Senior PySpark Data Engineer: ETL & Scalable Pipelines

Covetus • Irving (TX)

On-site
USD 100,000 - 130,000
Senior PySpark & Data Lakehouse Engineer
Senior PySpark & Data Lakehouse Engineer

Tieto • Irving (TX)

On-site
USD 140,000 - 190,000
Senior PySpark Data Engineer
Senior PySpark Data Engineer

Covetus • Irving (TX)

On-site
USD 100,000 - 130,000
Data Engineer
Data Engineer

The Value Maximizer • United States

On-site
USD 90,000 - 120,000
Data Engineer: Scalable Pipelines with Spark & Hive
Data Engineer: Scalable Pipelines with Spark & Hive

Tata Consultancy Services • Irving (TX)

On-site
USD 125,000 - 140,000
Senior Data Pipeline Engineer — Python & PySpark
Senior Data Pipeline Engineer — Python & PySpark

NTT DATA North America • Irving (TX)

On-site
USD 100,000 - 130,000
Senior Data Engineer — Spark, Python & Scala
Senior Data Engineer — Spark, Python & Scala

LTM • Irving (TX)

On-site
USD 100,000 - 130,000
Comprehensive Medical Plan Covering Medical, Dental, Vision
Short Term and Long-Term Disability Coverage
401(k) Plan with Company match
+2
AWS PySpark Data Engineer: Scalable Data Pipelines
AWS PySpark Data Engineer: Scalable Data Pipelines

LTM • Irving (TX)

On-site
USD 120,000 - 180,000
Medical plan
Disability coverage
401(k) match
+3