Senior Data Engineer (Spark & Python Specialist)

Hire Resolve

Johannesburg

On-site

ZAR 279,000 - 558,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Benefits offered by this job

Competitive salary based on experience

Job summary

A leading data collaboration platform in Johannesburg is seeking a Senior Cloud Data Engineer to lead the optimization of high-performance data processing engines. The ideal candidate will have strong expertise in Spark and Python, with a solid understanding of the Azure ecosystem, Delta Lake, and best practices in security and governance. This role offers competitive compensation based on experience. Interested candidates should apply by emailing their CV.

Qualifications

  • 6+ years of Spark/PySpark experience with optimization expertise.
  • Proficiency in building reusable Python libraries and automated testing.
  • Strong experience with Azure tools and services.

Responsibilities

  • Act as the internal SME for Spark internals and manage performance.
  • Build cloud-agnostic pipelines using Python and Delta Lake.
  • Migrate complex SQL-based ETL into modular Python libraries.
  • Manage Medallion Architecture ensuring optimal storage performance.
  • Support transition to code-centric orchestration patterns.

Skills

Spark/PySpark experience
Production-grade Python
Azure Synapse
Delta Lake
Docker
T-SQL
Security & Governance

Education

Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related field

Job description

An award-winning privacy-preserving data collaboration platform that enables companies to analyze and collaborate on consumer data to gain insights, build predictive models, and monetize data without sharing the raw data or compromising consumer privacy, is seeking a Senior Cloud Data Engineer who will be a lead technical contributor responsible for building and optimizing high-performance data processing engines.

Responsibilities
  • Spark Optimization: Act as the internal SME for Spark internals; manage memory, shuffle tuning, and partitioning for cost-effective performance.

  • Cloud-Agnostic Development: Build pipelines using Python and Delta Lake, decoupling code from specific cloud providers and reducing reliance on GUI tools (e.g., ADF).

  • Refactoring & Modernization: Migrate complex SQL-based ETL into modular, testable, and maintainable Python libraries.

  • Lakehouse Engineering: Manage Medallion Architecture (Bronze/Silver/Gold) using Delta Lake, ensuring storage performance via Z-Ordering and Vacuuming.

  • Code-First Orchestration: Support the transition to code-centric patterns (Airflow, Dagster) to prioritize portability.

  • Technical Excellence: Lead code reviews, mentor junior engineers, and implement automated testing frameworks (Pytest).

Minimum Requirements
  • Education: Bachelor’s degree in Computer Science, Information Systems, Engineering, or a related field.
  • Spark Mastery: 6+ years of Spark/PySpark experience; expert ability to diagnose bottlenecks via Spark UI and optimize complex DAGs.

  • Advanced Python: Proficiency in production-grade Python, including building reusable libraries and automated testing.

  • Azure Ecosystem: Strong experience with Azure Synapse, Dedicated SQL Pools, and Data Factory.

  • Modern Data Stack: Hands-on experience with Delta Lake, Parquet, and containerization (Docker).

  • Migration Skills: Solid T-SQL skills to interpret and migrate legacy logic into Python-centric environments.

  • Security & Governance: Proven ability to implement high levels of security and compliance across data processes.

Benefits
  • Competitive salary based on experience (salary can potentially be more based on experience/skills)

IFyou meet the above requirements and want to make a career-changing move, apply today by emailing your CV to itcareers@hireresolve.za.com

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Cloud Data Engineer - Remote
Senior Cloud Data Engineer - Remote

Hire Resolve • Cape Town

Remote
ZAR 700,000 - 950,000
Senior Data Science Consultant
Senior Data Science Consultant

Hire Resolve • Johannesburg

On-site
ZAR 1,200,000 - 1,800,000
Competitive salary based on experience
Senior Cloud Data Engineer — Spark & Python Lakehouse Lead
Senior Cloud Data Engineer — Spark & Python Lakehouse Lead

Hire Resolve • Johannesburg

On-site
ZAR 279,000 - 558,000
Competitive salary based on experience
Senior Data Solutions Engineer
Senior Data Solutions Engineer

Future Fit • Johannesburg

On-site
ZAR 900,000 - 1,500,000
Competitive compensation package
Twice-yearly salary increases
Employee wellness programs
+1
Senior Data Engineer
Senior Data Engineer

AES Global • Johannesburg

Hybrid
ZAR 900,000 - 1,500,000
Lead Data Scientist
Lead Data Scientist

MSP Staffing (PTY) LTD • Johannesburg

On-site
ZAR 1,200,000 - 1,600,000
Lead Data Scientist
Lead Data Scientist

MSP Staffing (PTY) LTD • Cape Town

On-site
ZAR 1,200,000 - 1,800,000
Lead Data Scientist
Lead Data Scientist

MSP Staffing (PTY) LTD • Durban

On-site
ZAR 1,200,000 - 2,000,000
Senior Data Engineer - Big Data Technologies
Senior Data Engineer - Big Data Technologies

Placements24 • Randburg

Hybrid
ZAR 900,000 - 1,300,000
Competitive salary and annual bonus
Medical, dental, and vision benefits
Retirement savings plan with employer
+2
Data Engineer (Analytics & Data Platform)
Data Engineer (Analytics & Data Platform)

ATS Client • Cape Town

On-site
ZAR 1,200,000 - 1,800,000
Training budget
Flexible working arrangements
Career development