Get more replies from employers
Send a job-specific resume in minutes.
GMG is seeking a highly skilled Data Engineer with a strong focus on AWS and Databricks to design, build, and maintain scalable data pipelines. The role involves ingestion of data from multiple sources, including Google Analytics, and optimizing performance and cost across AWS services and Databricks environments.
The candidate should have hands-on experience with Glue, PySpark, SQL, Athena, Lambda, SNS, and S3, along with CI/CD practices and governance standards.
We are seeking a highly skilled Data Engineer specializing in AWS and Databricks. The ideal candidate will design, build, and maintain scalable data pipelines, ensuring efficient data ingestion, processing, and integration from multiple sources—including Google Analytics event data. This role requires deep expertise in AWS Glue, Lambda, Athena, Redshift, Databricks, PySpark, and SQL, alongside strong performance tuning, data security, and cost optimization skills. Candidates with prior experience working in the retail domain will be strongly preferred. Cost optimization skills are essential.
The incumbent holds no direct supervisory responsibilities but is expected to engage effectively within their function and collaboratively across cross-functional teams.
This role contributes meaningfully, whether through operational contributions and/or by offering specialized expertise, guidance, and support, to ensure alignment with either functional and/or strategic organizational goals and objectives.
Knowledge of Glue, PySpark, SQL, Athena, Lambda, SNS, S3
Knowledge of Databricks: Cluster setup, Notebooks, Libraries, CI/CD, Optimization
Data Processing: Event stream ingestion and batch processing
Testing: Writing unit test cases and integration tests
Security & Governance: AWS/Databricks governance standards and best practices
Performance Optimization: Query tuning, cluster performance improvements, cost reduction
Strong problem-solving and analytical skills
Ability to work in a fast-paced, cloud-based data environment
Excellent collaboration and communication skills
Strong attention to detail and commitment to best practices
Minimum 6 experience in Data engineering (Core development/design), in which 3+ years on AWS with strong hands on (AWS glue, pyspark, SQL, Athena, lambda, SNS, S3) and 2+ year on Databricks.