AWS PySpark Data Engineer - Python Pipelines

Polarits

Newark (NJ)

On-site

USD 140,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Polarits in Newark, New Jersey, is seeking an experienced Python developer with strong AWS and PySpark expertise to join our data engineering team. You will design, build, and maintain scalable data pipelines and processing workflows in cloud environments.

The ideal candidate has 10+ years in software development with Python, hands-on PySpark, and a solid AWS background (S3, Glue, EMR, Redshift). You will collaborate across teams to deliver reliable, well-documented data solutions while

Qualifications

  • Bachelor’s degree in Computer Science, Data Engineering, or a related field.
  • 10+ years of experience in software development with a strong focus on Python.
  • Hands-on experience with PySpark for distributed data processing.
  • Solid understanding of AWS cloud services such as S3, Glue, Lambda, EMR, Redshift, and Athena.
  • Strong experience in ETL development and data pipeline orchestration.
  • Familiarity with SQL and relational/non-relational databases.
  • Excellent analytical, debugging, and communication skills.

Responsibilities

  • Design, develop, and maintain data pipelines and ETL workflows using Python, PySpark, and AWS services.
  • Build and optimize large-scale data processing and data transformation solutions.
  • Integrate various data sources and ensure data quality, performance, and reliability.
  • Collaborate with data engineers, analysts, and architects to deliver end-to-end data solutions.
  • Implement best practices for code optimization, error handling, and data validation.
  • Participate in code reviews, documentation, and deployment automation.
  • Ensure adherence to data security and compliance standards.

Skills

Python
PySpark
AWS
ETL
SQL
Distributed processing
Communication
Debugging

Education

Bachelor's degree in Computer Science or related field

Tools

Airflow
Databricks
Git
Docker
Kubernetes

Job description

Polarits in Newark, New Jersey, is seeking an experienced Python developer with strong AWS and PySpark expertise to join our data engineering team. You will design, build, and maintain scalable data pipelines and processing workflows in cloud environments.

The ideal candidate has 10+ years in software development with Python, hands-on PySpark, and a solid AWS background (S3, Glue, EMR, Redshift). You will collaborate across teams to deliver reliable, well-documented data solutions while

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AWS Python Developer with Pyspark Newark, NJ, New Jersey
AWS Python Developer with Pyspark Newark, NJ, New Jersey

Polarits • Newark (NJ)

On-site
USD 140,000 - 190,000
Senior AWS-Python Data Engineer (PySpark) - Hybrid NJ
Senior AWS-Python Data Engineer (PySpark) - Hybrid NJ

Polarits • Newark (NJ)

Hybrid
USD 140,000 - 180,000
AWS PySpark Data Engineer: Scalable Data Pipelines
AWS PySpark Data Engineer: Scalable Data Pipelines

LTM • Irving (TX)

On-site
USD 120,000 - 180,000
Medical plan
Disability coverage
401(k) match
+3
AWS Python Developer with Pyspark
AWS Python Developer with Pyspark

Polarits • Newark (NJ)

Hybrid
USD 140,000 - 180,000
Senior PySpark Data Engineer - AWS Data Pipelines
Senior PySpark Data Engineer - AWS Data Pipelines

Highbrow LLC • Charlotte (NC)

On-site
USD 110,000 - 160,000
Data Engineer - PySpark & AWS Cloud Pipelines
Data Engineer - PySpark & AWS Cloud Pipelines

SDLC Technologies • Charlotte (NC)

On-site
USD 90,000 - 150,000
Senior Data Engineer - PySpark & Python Pipelines
Senior Data Engineer - PySpark & Python Pipelines

CIS TECHNOLOGIES INC • Charlotte (NC)

On-site
USD 120,000 - 160,000
Data Engineer: Python Pipelines & PySpark Expert
Data Engineer: Python Pipelines & PySpark Expert

Insight Global • New York (NY)

On-site
USD 120,000 - 160,000
Senior AWS Data Engineer - Batch Pipelines & Spark
Senior AWS Data Engineer - Batch Pipelines & Spark

Polarits • Greenwood Village (CO)

On-site
USD 120,000 - 180,000
Data Engineer: Databricks, Python & Data Pipelines
Data Engineer: Databricks, Python & Data Pipelines

Cloud Analytics Technologies, LLC • Jersey City (NJ)

On-site