AWS Data Engineer (Databricks, AWS, Python)

Josh Pros LLC

Houston (TX)

On-site

USD 83,000 - 131,000

Full time

9 days ago
Application generator

Get a reply from this recruiter — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Job summary

Josh Pros LLC is seeking an AWS Data Engineer for an onsite contract in Houston, TX. You will design, build, and operate large-scale data pipelines on AWS and Databricks, turning raw data into reliable, well-modelled datasets powering analytics.

Responsibilities include building batch and streaming ETL/ELT pipelines, developing Delta Lake tables, and tuning Spark/SQL for performance and cost, with a focus on data quality, observability, and CI/CD deployments.

Qualifications

  • Must have AWS, Databricks, Python, PySpark and strong Spark performance tuning experience.
  • ,

Responsibilities

  • Build and maintain batch and streaming ETL/ELT pipelines using Databricks, PySpark, and Python.
  • Develop and optimise Delta Lake tables and medallion architectures.
  • Engineer data solutions on AWS using S3, Glue, EMR, Lambda, Redshift, Athena, Step Functions and Kinesis.
  • Model data for analytics and reporting; tune Spark jobs and SQL for performance and cost.
  • Implement data quality checks, monitoring, alerting, and lineage across pipelines.
  • Automate deployments with CI/CD and infrastructure-as-code; manage workflow orchestration.
  • Partner with analytics, product, and business teams to deliver production-grade data products.

Skills

AWS
Databricks
Python
PySpark
Spark tuning
SQL
Delta Lake
Unity Catalog
Medallion architecture
Kinesis
dbt
Airflow
Terraform
CloudFormation
CI/CD
Git
Observability
Data quality

Education

Bachelor's degree in Computer Science or related field

Tools

Databricks
Snowflake
Airflow
Terraform
CloudFormation
S3
Glue
EMR

Job description

We are seeking an AWS Data Engineer for an onsite contract (W-2) engagement in Houston, TX. You will design, build, and operate large-scale data pipelines on AWS and Databricks, turning raw source data into reliable, well-modelled datasets that power analytics and downstream applications.

Responsibilities
  • Build and maintain batch and streaming ETL/ELT pipelines using Databricks, PySpark, and Python
  • Develop and optimise Delta Lake tables and medallion (bronze/silver/gold) architectures
  • Engineer data solutions on AWS using S3, Glue, EMR, Lambda, Redshift, Athena, Step Functions, and Kinesis
  • Model data for analytics and reporting; tune Spark jobs and SQL for performance and cost
  • Implement data quality checks, monitoring, alerting, and lineage across pipelines
  • Automate deployments with CI/CD and infrastructure-as-code; manage workflow orchestration
  • Partner with analytics, product, and business teams to deliver production-grade data products
Work mode

This is a 100% onsite role in Houston, TX — no remote or hybrid option.

Employment

Contract on our W-2 only. Candidates must hold independent work authorisation (no third-party employer or sponsorship arrangement). Rate depends on experience.

Requirements
  • Must have: AWS
  • Must have: Databricks
  • Must have: Python
  • Strong PySpark and Spark performance tuning experience
  • Advanced SQL and dimensional/analytical data modelling
  • Delta Lake
  • Unity Catalog
  • and medallion architecture experience
  • AWS data services: S3
  • Glue
  • EMR
  • Lambda
  • Redshift
  • Athena
  • Step Functions
  • Kinesis
  • Workflow orchestration (Airflow
  • Databricks Workflows
  • or similar)
  • CI/CD
  • Git
  • and infrastructure-as-code (Terraform or CloudFormation)
  • Data quality
  • observability
  • and pipeline monitoring practices
  • Good to have: Snowflake
  • Kafka
  • dbt
  • streaming/real-time pipelines
  • Bachelor's degree in Computer Science
  • Engineering
  • or a related field
  • Work mode: onsite in Houston
  • TX (mandatory)
  • Employment: our W-2 only — independent work authorisation required
  • no third-party or sponsorship
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

AWS Data Engineer
AWS Data Engineer

Qode Page • United States

On-site
USD 140,000 - 180,000
Databricks Data Engineer
Databricks Data Engineer

VOLTO Consulting • Irving (TX)

On-site
USD 120,000 - 150,000
Azure Databricks Engineer
Azure Databricks Engineer

ZEUS SOLUTIONS INC • Houston (TX)

On-site
USD 140,000 - 180,000
AWS Data Engineer - w2
AWS Data Engineer - w2

SIDRAM TECHNOLOGIES • Reston (VA)

On-site
USD 60,000
Data Engineer - AWS/Databricks - Mid Level
Data Engineer - AWS/Databricks - Mid Level

Acuity, Inc. • Reston (VA)

On-site
USD 100,000 - 130,000
AWS Data Engineer
AWS Data Engineer

Qode • United States

Remote
USD 120,000 - 150,000
Senior Data Engineer on-site)
Senior Data Engineer on-site)

Ziosk • Dallas (TX)

On-site
USD 140,000 - 190,000
AWS Cloud Data Engineer
AWS Cloud Data Engineer

Eliassen Group • Nashville (TN)

On-site
Confidential
Medical, Dental, and Vision benefits
401k with company matching
Life insurance
AWS Cloud Engineer
AWS Cloud Engineer

Eliassen Group • Nashville (TN)

On-site
USD 157,597,000 - 171,924,000
Medical benefits
Dental benefits
Vision benefits
+2
Software Engineer III - Python, Databricks and AWS
Software Engineer III - Python, Databricks and AWS

hackajob • Jersey City (NJ)

On-site
USD 120,000 - 180,000
Health care coverage
On-site health and wellness centers
Retirement savings plan
+4