A technology company in Kennesaw, GA, seeks a Data Engineer with 6-8 years of experience in developing data solutions using Python and the Spark framework. Ideal candidates will design scalable data processes, perform root cause analysis, and resolve data-related issues. Proficiency in SQL and cloud knowledge, particularly Azure, is essential. Hands-on experience with Azure Synapse and Data Factory is a plus. This role involves close collaboration with stakeholders and IT teams.
Qualifications
6 - 8 years of experience in developing data solutions using Python and Spark.
Ability to perform root cause analysis and identify performance bottlenecks.
Hands-on knowledge of designing and developing data platforms in PySpark.
Responsibilities
Design and build new and scalable data processes.
Collaborate with stakeholders to resolve data-related issues.
Perform data analysis to troubleshoot issues.
Skills
Data Engineering
SQL
Cloud knowledge (Azure)
PySpark
Performance Bottleneck Analysis
Tools
Azure Synapse
Azure Data Factory
Job description
6 - 8 years’ of experience on developing data solutions in Python using Spark framework
Job Description
Designs, modifies, and builds new and scalable data processes.
Ability to perform root cause analysis and identify performance bottlenecks in Spark Jobs.
Expert in Data Engineering and building data pipelines, implementing Algorithms in a distributed environment.
Ability to design and develop parallel processing data platform in PySpark.
Performs data analysis required to troubleshoot data related issues and assist in the resolution of data issues.
Strong Proficiency in SQL.
Cloud knowledge especially Azure.
Collaborates with stakeholders, IT, database engineers and other scientists.
Hands-on knowledge in Azure Synapse and Azure Data Factory is a plus.