Data Engineer/Python Developer

TechDigital Group

Minnesota

On-site

USD 80,000 - 120,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

An innovative firm is seeking a skilled Data Engineer to develop and optimize ETL/ELT pipelines using PySpark and SQL. In this role, you will work with both structured and unstructured data to build scalable data solutions, ensuring data quality and performance optimization. Collaborating closely with Data Scientists and Analysts, you will design data models and implement robust data workflows on cloud platforms. This position offers a fantastic opportunity to contribute to cutting-edge data projects and enhance your skills in a dynamic and supportive environment.

Qualifications

  • Strong experience in Python and PySpark for data engineering tasks.
  • Deep understanding of SQL and ETL/ELT development using Spark.

Responsibilities

  • Develop and maintain ETL/ELT pipelines using PySpark and SQL.
  • Collaborate with Data Scientists and Analysts to integrate data workflows.

Skills

Python
PySpark
SQL
ETL/ELT Development
Cloud Data Services
Data Warehousing
Orchestration Tools
Version Control (Git)

Tools

AWS Glue
Databricks
Azure Synapse
GCP BigQuery
Airflow
Apache Oozie
Snowflake
Redshift

Job description

Job Responsibilities:
  1. Develop, optimize, and maintain ETL/ELT pipelines using PySpark and SQL.
  2. Work with structured and unstructured data to build scalable data solutions.
  3. Write efficient and scalable PySpark scripts for data transformation and processing.
  4. Optimize SQL queries, stored procedures, and indexing strategies to enhance performance.
  5. Design and implement data models, schemas, and partitioning strategies for large-scale datasets.
  6. Collaborate with Data Scientists, Analysts, and other Engineers to integrate data workflows.
  7. Ensure data quality, validation, and consistency in data pipelines.
  8. Implement error handling, logging, and monitoring for data pipelines.
  9. Work with cloud platforms (AWS, Azure, or GCP) for data processing and storage.
  10. Optimize data pipelines for cost efficiency and performance.
Technical Skills Required:
  1. Strong experience in Python for data engineering tasks.
  2. Proficiency in PySpark for large-scale data processing.
  3. Deep understanding of SQL (Joins, Window Functions, CTEs, Query Optimization).
  4. Experience in ETL/ELT development using Spark and SQL.
  5. Experience with cloud data services (AWS Glue, Databricks, Azure Synapse, GCP BigQuery).
  6. Familiarity with orchestration tools (Airflow, Apache Oozie).
  7. Experience with data warehousing (Snowflake, Redshift, BigQuery).
  8. Understanding of performance tuning in PySpark and SQL.
  9. Familiarity with version control (Git) and CI/CD pipelines.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

SQL Developer
SQL Developer

TechDigital Group • United States

On-site
USD 80,000 - 120,000
Data Engineer
Data Engineer

The Value Maximizer • United States

On-site
USD 90,000 - 120,000
Data Engineer
Data Engineer

Jobtailor • Kentucky

On-site
USD 110,000 - 140,000
PySpark Developer
PySpark Developer

Inizio Partners Corp • Hartford (CT)

On-site
USD 90,000 - 120,000
Data Engineer
Data Engineer

SDLC Technologies • Charlotte (NC)

On-site
USD 90,000 - 150,000
Data Engineer ETL
Data Engineer ETL

Compunnel, Inc. • Durham (NC)

On-site
USD 100,000 - 130,000
Data Engineer
Data Engineer

TMV Global Inc • Virginia (MN)

On-site
USD 95,000 - 120,000
Data Engineer
Data Engineer

ATC • New York (NY)

On-site
USD 140,000 - 180,000
Data Engineer
Data Engineer

Compunnel, Inc. • Charlotte (NC)

On-site
USD 90,000 - 120,000
Senior PySpark Data Engineer
Senior PySpark Data Engineer

Covetus • Irving (TX)

On-site
USD 100,000 - 130,000