AWS PySpark Data Engineer: Scalable Data Pipelines

LTM

Irving (TX)

On-site

USD 120,000 - 180,000

Full time

14 days+
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Benefits offered by this job

Medical plan
Disability coverage
401(k) match
Life insurance
Paid time off
Parental leave

Job summary

LTIMindtree in Irving, Texas seeks an AWS PySpark Developer to design, build, and optimize scalable data solutions on the AWS ecosystem. The role focuses on PySpark data processing, Iceberg/Hive data warehousing, and secure, cost-efficient pipelines.

You will deploy data solutions using S3, EMR, Glue, Athena, Lambda, and Redshift, while ensuring strong governance, performance, and collaboration with data scientists and engineers. This is a full-time, on-site position in the US.

Qualifications

  • AWS certification(s) required or in progress.
  • Hands-on with core AWS data services (S3, EC2, EMR, Glue, Athena, Lambda, Redshift, Kinesis).
  • Strong PySpark data processing and open table formats like Iceberg.
  • Experience with Hive-based data warehousing and data modeling.

Responsibilities

  • Design, implement, and deploy scalable data solutions on AWS.
  • Develop robust PySpark ETL/ELT pipelines for ingestion and loading.
  • Build and optimize data pipelines using Iceberg and Hive formats.
  • Design data warehousing schemas and optimize queries for analytics.
  • Configure secure cloud infrastructure, including VPCs and security groups.
  • Operate containerized workloads on EKS and manage Spark/Hive jobs.
  • Tune performance and manage cost of big data applications on AWS.
  • Implement data governance, security, and access controls in AWS.

Skills

AWS
PySpark
Data warehousing
Iceberg
Hive
ETL/ELT
Python
SQL
Boto3
Kubernetes

Tools

S3
EC2
EMR
Glue
Athena
Lambda
Redshift
Kinesis
Iceberg

Job description

LTIMindtree in Irving, Texas seeks an AWS PySpark Developer to design, build, and optimize scalable data solutions on the AWS ecosystem. The role focuses on PySpark data processing, Iceberg/Hive data warehousing, and secure, cost-efficient pipelines.

You will deploy data solutions using S3, EMR, Glue, Athena, Lambda, and Redshift, while ensuring strong governance, performance, and collaboration with data scientists and engineers. This is a full-time, on-site position in the US.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior AWS Data Engineer: PySpark & Iceberg Architect
Senior AWS Data Engineer: PySpark & Iceberg Architect

LTM • Irving (TX)

On-site
USD 120,000 - 170,000
Medical Plan
Disability Insurance
401(k) with match
+3
Senior Data Engineer - AWS PySpark & Iceberg Pipelines
Senior Data Engineer - AWS PySpark & Iceberg Pipelines

LTM • Tampa (FL)

On-site
USD 95,000 - 140,000
Comprehensive Medical Plan
Disability Coverage
401(k) Plan with Company match
+3
Senior AWS Data Engineer: PySpark, Iceberg & Hive
Senior AWS Data Engineer: PySpark, Iceberg & Hive

LTM • Tampa (FL)

On-site
USD 140,000 - 190,000
Comprehensive Medical Plan
401(k) Plan with Company match
Paid Paternity and Maternity Leave
+2
Senior Data Engineering Lead - Spark & Python
Senior Data Engineering Lead - Spark & Python

LTM • Irving (TX)

On-site
USD 125,000 - 180,000
Comprehensive Medical Plan
Disability Coverage
401(k) Plan with Company match
+3
Data Engineer: Scalable Pipelines with Spark & Hive
Data Engineer: Scalable Pipelines with Spark & Hive

Tata Consultancy Services • Irving (TX)

On-site
USD 125,000 - 140,000
Spark Data Engineer (PySpark) - Cloud ETL & Pipelines
Spark Data Engineer (PySpark) - Cloud ETL & Pipelines

Tata Consultancy Services • Irving (TX)

On-site
USD 90,000 - 120,000
Discretionary Annual Incentive
Medical Coverage (Health, Dental & Vis
401K Plan
+1
Senior PySpark Data Engineer: ETL & Scalable Pipelines
Senior PySpark Data Engineer: ETL & Scalable Pipelines

Covetus • Irving (TX)

On-site
USD 100,000 - 130,000
Cloud Data Engineer: Spark & Hive Data Pipelines
Cloud Data Engineer: Spark & Hive Data Pipelines

Tata Consultancy Services • Irving (TX)

On-site
USD 90,000 - 110,000
Lead PySpark Developer — Azure Data Pipelines & AKS
Lead PySpark Developer — Azure Data Pipelines & AKS

Wipro Technologies • Minneapolis (MN)

Hybrid
USD 80,000 - 158,000
Medical and dental insurance
Paid time off
Disability coverage
Data Engineer - Spark, Airflow & ETL Pipelines
Data Engineer - Spark, Airflow & ETL Pipelines

TP-Link • Irvine (CA)

On-site
USD 100,000 - 120,000
Free snacks and drinks
Fully paid medical, dental, and vision
401K contribution
+3