Senior Data Engineer (Spark & Python Specialist)

Hire Resolve

Johannesburg

On-site

ZAR 279,000 - 558,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Competitive salary based on experience

Job summary

A leading data collaboration platform in Johannesburg is seeking a Senior Cloud Data Engineer to lead the optimization of high-performance data processing engines. The ideal candidate will have strong expertise in Spark and Python, with a solid understanding of the Azure ecosystem, Delta Lake, and best practices in security and governance. This role offers competitive compensation based on experience. Interested candidates should apply by emailing their CV.

Qualifications

  • 6+ years of Spark/PySpark experience with optimization expertise.
  • Proficiency in building reusable Python libraries and automated testing.
  • Strong experience with Azure tools and services.

Responsibilities

  • Act as the internal SME for Spark internals and manage performance.
  • Build cloud-agnostic pipelines using Python and Delta Lake.
  • Migrate complex SQL-based ETL into modular Python libraries.
  • Manage Medallion Architecture ensuring optimal storage performance.
  • Support transition to code-centric orchestration patterns.

Skills

Spark/PySpark experience
Production-grade Python
Azure Synapse
Delta Lake
Docker
T-SQL
Security & Governance

Education

Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related field

Job description

An award-winning privacy-preserving data collaboration platform that enables companies to analyze and collaborate on consumer data to gain insights, build predictive models, and monetize data without sharing the raw data or compromising consumer privacy, is seeking a Senior Cloud Data Engineer who will be a lead technical contributor responsible for building and optimizing high-performance data processing engines.

Responsibilities
  • Spark Optimization: Act as the internal SME for Spark internals; manage memory, shuffle tuning, and partitioning for cost-effective performance.

  • Cloud-Agnostic Development: Build pipelines using Python and Delta Lake, decoupling code from specific cloud providers and reducing reliance on GUI tools (e.g., ADF).

  • Refactoring & Modernization: Migrate complex SQL-based ETL into modular, testable, and maintainable Python libraries.

  • Lakehouse Engineering: Manage Medallion Architecture (Bronze/Silver/Gold) using Delta Lake, ensuring storage performance via Z-Ordering and Vacuuming.

  • Code-First Orchestration: Support the transition to code-centric patterns (Airflow, Dagster) to prioritize portability.

  • Technical Excellence: Lead code reviews, mentor junior engineers, and implement automated testing frameworks (Pytest).

Minimum Requirements
  • Education: Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related field.
  • Spark Mastery: 6+ years of Spark/PySpark experience; expert ability to diagnose bottlenecks via Spark UI and optimize complex DAGs.

  • Advanced Python: Proficiency in production-grade Python, including building reusable libraries and automated testing.

  • Azure Ecosystem: Strong experience with Azure Synapse, Dedicated SQL Pools, and Data Factory.

  • Modern Data Stack: Hands-on experience with Delta Lake, Parquet, and containerization (Docker).

  • Migration Skills: Solid T-SQL skills to interpret and migrate legacy logic into Python-centric environments.

  • Security & Governance: Proven ability to implement high levels of security and compliance across data processes.

Benefits
  • Competitive salary based on experience (salary can potentially be more based on experience/skills)

IFyou meet the above requirements and want to make a career-changing move, apply today by emailing your CV to itcareers@hireresolve.za.com

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer (Spark & Python Specialist)
Senior Data Engineer (Spark & Python Specialist)

WatersEdge Solutions • Cape Town

Remote
ZAR 1,000,000 - 1,400,000
Flexible working arrangements
Wellness support and home office reimbursement
Continuous learning opportunities
+2
Senior Cloud Data Engineer - Remote
Senior Cloud Data Engineer - Remote

Hire Resolve • Cape Town

Remote
ZAR 700,000 - 950,000
Senior Data Science Consultant
Senior Data Science Consultant

Hire Resolve • Johannesburg

On-site
ZAR 1,200,000 - 1,800,000
Competitive salary based on experience
Senior Cloud Data Engineer — Spark & Python Lakehouse Lead
Senior Cloud Data Engineer — Spark & Python Lakehouse Lead

Hire Resolve • Johannesburg

On-site
Competitive salary based on experience
Senior Data Solutions Engineer
Senior Data Solutions Engineer

Future Fit • Johannesburg

Hybrid
ZAR 900,000 - 1,500,000
Competitive compensation package
Twice-yearly salary increases
Employee wellness programs
+1
Senior Data Engineer
Senior Data Engineer

CloudDevs • South Africa

Remote
ZAR 600,000 - 800,000
Flexible working hours
Competitive compensation
Company-sponsored learning and certifications
+3
Cloud Data Engineer (12-month contract)
Cloud Data Engineer (12-month contract)

DeARX • Sandton

On-site
ZAR 900,000 - 1,300,000
Data Engineer (Analytics & Data Platform)
Data Engineer (Analytics & Data Platform)

ATS Client • Cape Town

Hybrid
ZAR 1,200,000 - 1,800,000
Training budget
Flexible working arrangements
Career development
Data Engineer
Data Engineer

Calibrate People • Johannesburg

Hybrid
ZAR 600,000 - 800,000
Competitive salary
Data Engineering Lead
Data Engineering Lead

Blue Pearl PTY • Johannesburg

On-site
ZAR 1,200,000 - 1,900,000