PySpark Consultant

VDart

Irving (TX)

On-site

USD 120,000 - 150,000

Full time

2 days ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

VDart is seeking a PySpark Consultant with minimum 6+ years of experience for a full-time role in Irving, TX. The role focuses on building and maintaining PySpark-based ETL pipelines, ensuring data quality, and optimizing distributed data workflows.

Strong coding skills in Python and SQL, plus experience with Hadoop, Hive, and Kafka, are essential to succeed. You will lead data processing efforts, document code lineage, and collaborate with data ingestion and analytics teams while adhering to

Qualifications

  • +4+ years of experience in big data development, Hadoop, Hive & Spark framework.
  • Strong Python, PySpark Development and SQL knowledge.
  • Certification in big data or cloud technologies is preferred.
  • Experience with SAS is a plus.

Responsibilities

  • Design, develop and optimize PySpark ETL pipelines.
  • Ensure data quality and integrity across data processing workflows.
  • Troubleshoot and resolve issues in PySpark applications and workflows.
  • Understand source data, dependencies and data flow from converted PySpark code.
  • Integrate PySpark code with frameworks such as Ingestion Framework, DataLens, etc.
  • Ensure compliance with data security, privacy regulations, and organizational standards.

Skills

PySpark
Python
SQL
Hadoop
Hive
Kafka
Data Warehousing
CI/CD
DevOps
Leadership
Communication

Tools

Hadoop
Hive
Kafka

Job description

Role : PySpark Consultant ( Minimum 6+ years )
Location : Irving, TX
Job Type : Full-time only.
Job Description :
  • Experience with big data processing and distributed computing systems like Spark.
  • Implement ETL pipelines and data transformation processes.
  • Ensure data quality and integrity in all data processing workflows.
  • Troubleshoot and resolve issues related to PySpark applications and workflows.
  • Understand source, dependencies and data flow from converted PySpark code.
  • Strong programming skills in Python and SQL.
  • Experience with big data technologies like Hadoop, Hive, and Kafka.
  • Understanding of data warehousing concepts and relational databases like SQL.
  • Demonstrate and document code lineage.
  • Integrate PySpark code with frameworks such as Ingestion Framework, DataLens, etc.,
  • Ensure compliance with data security, privacy regulations, and organizational standards.
  • Knowledge of CI/CD pipelines and DevOps practices.
  • Strong problem-solving and analytical skills.
  • Excellent communication and leadership abilities. Qualifications:
  • 4+ years of experience in big data development, Hadoop , Hive & Spark framework.
  • Good to have experience in SAS.
  • Strong Python, PySpark Development and SQL knowledge.
  • Certification in big data or cloud technologies is preferred.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

PySpark Consultant Contract C2C jobs Urgent Need
PySpark Consultant Contract C2C jobs Urgent Need

Tech Mirrors • Irving (TX)

On-site
USD 96,000 - 152,000
PySpark Consultant
PySpark Consultant

VDart Inc • Irving (TX)

On-site
USD 120,000 - 160,000
Senior PySpark Data Engineer — ETL & Big Data
Senior PySpark Data Engineer — ETL & Big Data

VDart • Irving (TX)

On-site
USD 120,000 - 150,000
Data Engineer
Data Engineer

The Value Maximizer • United States

On-site
USD 90,000 - 120,000
Senior PySpark Data Engineer | ETL & Data Quality
Senior PySpark Data Engineer | ETL & Data Quality

Tech Mirrors • Irving (TX)

On-site
USD 96,000 - 152,000
Hadoop developer
Hadoop developer

REALIGN LLC • Charlotte (NC)

On-site
USD 120,000 - 180,000
Pyspark Developer
Pyspark Developer

Tieto • Irving (TX)

On-site
USD 140,000 - 190,000
PySpark Data Engineer
PySpark Data Engineer

Tata Consultancy Services • Irving (TX)

On-site
USD 90,000 - 120,000
Discretionary Annual Incentive
Medical Coverage (Health, Dental & Vis
401K Plan
+1
PySpark Engineer - Big Data & Spark Expert
PySpark Engineer - Big Data & Spark Expert

Infosys Limited • Irving (TX)

On-site
USD 90,000 - 130,000
Long-term Disability
Health and Dependent Care Reimbursement Accounts
401(k) plan
Pyspark developer (Data engineer)
Pyspark developer (Data engineer)

MTK Technologies • Irving (TX)

On-site
USD 90,000 - 120,000