Cloud Data Engineer - Spark, PySpark & DataBricks

Highbrow LLC

Town of Texas (WI)

Hybrid

USD 120,000 - 160,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Highbrow LLC is seeking a Cloud Engineer to work with Austin-based leadership to design and implement data extraction and transformation pipelines using Spark, PySpark, and DataBricks.

The role involves writing complex SQL across Teradata, Hive, and Spark, building DevOps pipelines, and contributing to agile ceremonies. 3 days/week on-site at the Austin office with collaboration across teams.

Qualifications

  • Strong development experience in Spark, PySpark, Shell scripting, Teradata, and DataBricks.
  • Experience with Ab Initio is required.
  • Strong experience in writing complex and effective SQLs ( Teradata, Hive, Spark ) and Stored Procedures.
  • Proficiency in Shell Scripting for automation, system administration, and data processing tasks.
  • Excellent work experience on DataBricks as data warehouse/Data Lake implementations.
  • Experience in Agile and DevOps tools (Git, Jenkins, Artifactory), Unix/Linux Shell scripting (KSH) and basic Unix server administration.
  • Hands-on experience with Teradata data modeling, SQL development, performance tuning, and large-scale data management.
  • Experience in Jira and Confluence.

Responsibilities

  • Collaborate with leadership to understand requirements and design solutions.
  • Write product-grade code for data extraction, transformation, and loading using Spark/PySpark.
  • Model data and write performant queries using Spark SQL, Teradata SQL and Hive SQL.
  • Implement DevOps pipelines to deploy code artifacts to AWS/DataBricks.
  • Orchestrate DataBricks jobs and monitor production data pipelines.
  • Participate in agile ceremonies like sprint planning, grooming, daily scrums, and demos.

Skills

Spark/PySpark
Shell scripting
Teradata
DataBricks
Ab Initio
SQL (Teradata/Hive/Spark)
Unix/Linux Shell scripting (KSH)
DevOps practices
Git/Jenkins/Artifactory
Databricks Delta Lake

Tools

Jira
Confluence
Git
Jenkins
Artifactory
CA7 Scheduler
AWS (S3, EC2, Lambda, etc.)
DataBricks Platform
Unix Administration

Job description

Highbrow LLC is seeking a Cloud Engineer to work with Austin-based leadership to design and implement data extraction and transformation pipelines using Spark, PySpark, and DataBricks.

The role involves writing complex SQL across Teradata, Hive, and Spark, building DevOps pipelines, and contributing to agile ceremonies. 3 days/week on-site at the Austin office with collaboration across teams.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Spark & DataBricks Cloud Engineer (AWS, SQL, DevOps)
Spark & DataBricks Cloud Engineer (AWS, SQL, DevOps)

Highbrow LLC • Town of Texas (WI)

On-site
USD 120,000 - 160,000
Cloud Engineer
Cloud Engineer

Highbrow LLC • Town of Texas (WI)

Hybrid
USD 120,000 - 160,000
Cloud Engineer (Client: -)
Cloud Engineer (Client: -)

Highbrow LLC • Town of Texas (WI)

On-site
USD 120,000 - 160,000
Cloud Data Engineer: AWS Databricks & Spark ETL
Cloud Data Engineer: AWS Databricks & Spark ETL

TalentOla • Austin (TX)

On-site
USD 120,000 - 160,000
Senior Cloud Data Engineer: Databricks & Spark
Senior Cloud Data Engineer: Databricks & Spark

CGI Group, Inc. • Salt Lake City (UT)

On-site
USD 80,000 - 139,000
Competitive compensation
Comprehensive insurance options
401(k) matching contributions and the
+4
Databricks Data Engineer - Spark, ETL & Cloud Pipelines
Databricks Data Engineer - Spark, ETL & Cloud Pipelines

Smart IT Frame LLC • Reston (VA)

On-site
USD 90,000 - 120,000
Senior Data Engineer: PySpark, Databricks & Azure Pipelines
Senior Data Engineer: PySpark, Databricks & Azure Pipelines

BrickRed Systems • Frisco (TX)

On-site
USD 120,000 - 170,000
Senior Data Engineer — Python, Databricks on AWS
Senior Data Engineer — Python, Databricks on AWS

hackajob • Jersey City (NJ)

On-site
USD 120,000 - 180,000
Health care coverage
On-site health and wellness centers
Retirement savings plan
+4
Senior Data Engineer — Spark, Databricks & Cloud Pipelines
Senior Data Engineer — Spark, Databricks & Cloud Pipelines

ATC • New York (NY)

On-site
USD 140,000 - 180,000
Senior Data Engineer — Spark & Cloud Pipelines
Senior Data Engineer — Spark & Cloud Pipelines

Tata Consultancy Services • Irving (TX)

On-site
USD 100,000 - 130,000