Senior PySpark Data Engineer

Covetus

Irving (TX)

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Covetus, located in Irving, Texas, is seeking a Data Engineer who will be responsible for designing, developing, and maintaining data solutions in a Big Data environment using predominantly PySpark/Python. The role involves creating data pipelines, ensuring data quality, and implementing ETL processes across systems.

The ideal candidate should have over 8 years of experience in Hadoop and PySpark, with substantial exposure to AWS and Snowflake. Strong collaboration skills and problem-solving abilities are essential for success in this role.

Qualifications

  • 8+ years of professional experience in Hadoop, PySpark/Python development.
  • Proven expertise in PySpark and experience in handling huge volumes of data.
  • 3+ years of working experience in AWS, Databricks/Snowflake, Airflow.

Responsibilities

  • Design, develop, and maintain robust, scalable high-performance Data Pipelines using PySpark.
  • Create data pipelines, ensuring data quality, and implement ETL processes to migrate and deploy data across systems.
  • Migrate Ab Initio ETL Applications into PySpark based data pipelines.

Skills

Hadoop
PySpark/Python
AWS
Databricks/Snowflake
CI/CD
Git
Debugging
Communication

Job description

As a Data Engineer, you will be responsible for designing, developing, and maintaining data solutions for data generation, collection, and processing in Big Data environment using predominantly PySpark/Python. Your typical day will involve creating data pipelines, ensuring data quality, and implementing ETL processes to migrate and deploy data across systems using PySpark.

Responsibilities
  • Design, develop, and maintain robust, scalable high-performance Data Pipelines using PySpark.
  • Create data pipelines, ensuring data quality, and implement ETL processes to migrate and deploy data across systems.
  • Migrate Ab Initio ETL Applications into Pyspark based data pipelines.
  • Migrate On Prem workloads into Cloud (AWS, Databricks, Snowflake) based on use cases.
  • Collaborate with cross-functional teams to identify and resolve data-related issues.
  • Stay updated with the latest advancements in data engineering and integrate innovative approaches for sustained competitive advantage.
Qualifications
  • 8+ years of professional experience in Hadoop, PySpark/Python development.
  • Proven expertise in PySpark and experience in handling huge volume of data.
  • 3+ years of working experience in AWS, Databricks/Snowflake, Airflow.
  • Familiarity with CI/CD pipelines and version control systems (e.g., Git).
  • Strong debugging and problem-solving skills.
  • Excellent communication and collaboration skills.
Good to Have Skills
  • AWS EKS Experience, Dockers and Containers.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Engineer
Engineer

Tata Consultancy Services • Dallas (TX)

On-site
USD 100,000 - 105,000
PySpark Developer
PySpark Developer

Inizio Partners Corp • Hartford (CT)

On-site
USD 90,000 - 120,000
Data Engineer
Data Engineer

The Value Maximizer • United States

On-site
USD 90,000 - 120,000
Data Engineer
Data Engineer

SDLC Technologies • Charlotte (NC)

On-site
USD 90,000 - 150,000
Data Engineer
Data Engineer

Jobtailor • Kentucky

On-site
USD 110,000 - 140,000
Data Engineer/Python Developer
Data Engineer/Python Developer

TechDigital Group • Minnesota

On-site
USD 80,000 - 120,000
Pyspark Developer
Pyspark Developer

Tata Consultancy Services • Irving (TX)

On-site
USD 100,000 - 130,000
Senior Data Engineer – PySpark & Python
Senior Data Engineer – PySpark & Python

Highbrow LLC • Charlotte (NC)

On-site
USD 110,000 - 160,000
Data Engineer
Data Engineer

Siro Clinpharm • United States

Remote
USD 37,000 - 73,000
Data Engineer
Data Engineer

The Judge Group • New York (NY)

On-site
USD 150,000 - 190,000