Big Data Engineer: Spark, ETL, Data Quality

eNcloud

Boston, Northern (MA, KY)

Hybrid

USD 120,000 - 180,000

Full time

35 hours ago
Be an early applicant
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

eNcloud is seeking a data engineer in Boston to build data transformation and ingestion pipelines. You will develop ETL processes using Spark, Hive, HDFS and related tools, while handling large data volumes and ensuring data quality.

You will write high-performance SQL, shell scripts, and manage Hive tables and data structures. Collaboration with CI/CD and Bitbucket is required, with travel to client locations as needed.

Qualifications

  • Bachelor’s degree in Computer Science, IT, Engg or related
  • 5+ years of experience in data engineering or related role
  • Experience with big data tools and data ingestion/curation processes

Responsibilities

  • Build processes for data transformation, data structures, metadata, dependency and workload management
  • Develop data ingestion and curation using Spark, Hive, HDFS, Sqoop, HBase, Kerberos and Impala
  • Ingest large data volumes from multiple platforms and write high-performance ETL code with strong SQL and data quality focus
  • Write shell scripts, complex SQL queries, Hadoop commands and Git workflows
  • Create databases, schemas, Hive tables (External and Managed) with Orc/Parquet/Avro/Text formats
  • Monitor production job performance and advise infrastructure changes
  • Work with Bitbucket and CI/CD pipelines

Skills

Big data
Spark
Hive
SQL
ETL
Shell scripting
Git
CI/CD
Bitbucket
Hadoop

Education

Bachelor's degree

Tools

HDFS
Sqoop
HBase
Kerberos
Sentry
Impala

Job description

eNcloud is seeking a data engineer in Boston to build data transformation and ingestion pipelines. You will develop ETL processes using Spark, Hive, HDFS and related tools, while handling large data volumes and ensuring data quality.

You will write high-performance SQL, shell scripts, and manage Hive tables and data structures. Collaboration with CI/CD and Bitbucket is required, with travel to client locations as needed.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Big Data Engineer: Spark, ETL & Data Quality
Big Data Engineer: Spark, ETL & Data Quality

eNcloud Services LLC • Katy (TX)

On-site
USD 100,000 - 170,000
Big Data Developer - Spark, Hive & ETL Expert
Big Data Developer - Spark, Hive & ETL Expert

eNcloud • Katy (TX)

Hybrid
USD 95,000 - 140,000
Data Engineer ETL
Data Engineer ETL

Compunnel, Inc. • Durham (NC)

On-site
USD 100,000 - 130,000
Cloud Data Engineer: Spark & Hive Data Pipelines
Cloud Data Engineer: Spark & Hive Data Pipelines

Tata Consultancy Services • Irving (TX)

On-site
USD 90,000 - 110,000
Big Data Developer
Big Data Developer

eNcloud • Katy (TX)

Hybrid
USD 95,000 - 140,000
Senior Data Engineer: Spark & Cloud Pipelines
Senior Data Engineer: Spark & Cloud Pipelines

Tata Consultancy Services • Irving (TX)

On-site
USD 90,000 - 110,000
Discretionary Annual Incentive.
Comprehensive Medical Coverage
401K Plan
+1
Data Engineer: Snowflake, Spark & ETL for Analytics
Data Engineer: Snowflake, Spark & ETL for Analytics

Praise Tech Solutions • Northern (KY)

Hybrid
USD 120,000 - 160,000
Big Data Engineer: ETL, Spark & Data Pipelines
Big Data Engineer: ETL, Spark & Data Pipelines

Synechron • Charlotte (NC)

On-site
USD 100,000 - 110,000
Highly competitive compensation and benefits package
10 days of paid annual leave
Comprehensive insurance plan
+2
Production Data Engineer: Spark, ETL & Data Ops
Production Data Engineer: Spark, ETL & Data Ops

Tata Consultancy Services • Charlotte (NC)

On-site
USD 95,000 - 115,000
Data Engineer
Data Engineer

Curate Partners • Massachusetts

On-site
USD 90,000 - 130,000