Pyspark Architect

Avance Consulting

North Carolina

On-site

USD 100,000 - 130,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Avance Consulting is looking for a skilled technology consultant based in North Carolina. The role requires strong experience in the Cloudera Hadoop ecosystem, including Hive, Scala, and SPARK, as well as expertise in designing ETL/ELT frameworks for complex data warehouses.

The ideal candidate will have a background in software engineering and must be proficient in data virtualization and parallel data processing. This position offers the opportunity to solve complex problems in a dynamic environment.

Responsibilities

  • Design and implement ETL/ELT frameworks for data warehouses.
  • Work with large data sets; performance tuning and troubleshooting.
  • Develop complex data products in heterogeneous environments.

Skills

Technology consulting
Cloudera Hadoop ecosystem
ETL/ELT framework
Data virtualization
Microservice architecture

Tools

AWS
Azure
GCP
Python
Scala
Hadoop

Job description

  • Strong experience in technology consulting, enterprise and solutions architecture and architectural frameworks
  • Strong experience in Cloudera Hadoop ecosystem, i.e. Hadoop, Hive, Scala, SPARK, Sqoop, Flume, Kafka, Python, NoSQL database technologies
  • Good Understanding and knowledge of Hadoop architecture and various components such as high availability architecture, experience in workload management, scalability and distributed platform architecture
  • Experience with design and implementation of ETL/ELT framework for complex warehouses/marts. Knowledge of large data sets and experience with performance tuning and troubleshooting
  • Background in all aspects of software engineering with strong skills in parallel data processing, data flows, REST APIs, JSON, XML, and micro service architecture
Preferred Skill and Experience
  • Experience in working in any hyperscalers, AWS, Azure, GCP, Databricks and Snowflake
  • Experience in data virtualization/data federationfor creating data product (e.g. in Starburst) in heterogenous environment (Hadoop, RDBMS, NoSQL etc.)
  • Hands-on development mentality, with a willingness to troubleshoot and solve complex problems
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Pyspark Architect
Pyspark Architect

Avance Consulting • Charlotte (NC)

On-site
USD 100,000 - 130,000
Spark Engineer
Spark Engineer

Veriipro • Jacksonville (FL)

On-site
USD 90,000 - 130,000
Hadoop Developer
Hadoop Developer

Fixity Technologies • Charlotte (NC)

On-site
USD 120,000 - 180,000
Senior Technical Lead
Senior Technical Lead

Infinite Computer Solutions • Town of Texas (WI)

On-site
USD 120,000 - 180,000
Hadoop and PySpark Developer
Hadoop and PySpark Developer

Avance Consulting • Plano (TX)

On-site
USD 95,000 - 135,000
Pyspark Developer
Pyspark Developer

Tata Consultancy Services • Irving (TX)

On-site
USD 100,000 - 130,000
Java Spark Engineer
Java Spark Engineer

Veriipro • Berkeley Heights (NJ)

On-site
USD 140,000 - 190,000
Senior Software Engineer
Senior Software Engineer

Infinite Computer Solutions • Town of Texas (WI)

On-site
USD 110,000 - 140,000
Senior Data Engineer
Senior Data Engineer

Jobtailor • Santa Clara (CA)

On-site
USD 150,000 - 230,000
Data Engineer
Data Engineer

Infinite Computer Solutions • Town of Texas (WI)

On-site
USD 120,000 - 170,000