Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
WatersEdge Solutions is seeking a Lead Data Scientist for a senior, hands-on leadership role to own delivery, mentor team members and shape the evolution of a sophisticated data science function. You'll remain deeply technical, delivering production-grade Python solutions, building AI-powered pipelines and analytics.
The role involves Python, PySpark, Delta Lake, Azure Synapse, Trino, Power BI, Kubernetes and FastAPI, with remote work from South Africa and occasional travel to Johannesburg or
Location: South Africa | Remote
Employment Type: Permanent | Full-Time
Industry: Data Science | AI | Data Technology
WatersEdge Solutions is partnering with an innovative technology business to find an experienced Lead Data Scientist for a senior, hands-on leadership role. This is an opportunity for someone who wants to remain deeply involved in building data science solutions while also mentoring others, owning delivery and helping shape the direction of a sophisticated data science function.
The role sits at the heart of the product environment, where machine learning, privacy-preserving modelling, entity resolution, data linkage, analytics and AI-powered pipelines are central to delivering value to enterprise customers.
As Lead Data Scientist, you'll act as second-in-command to the Head of Data Science, combining technical leadership with hands-on delivery. You'll lead complex analytical work, take ownership of customer-facing analytical delivery and represent the Data Science function when required.
Importantly, this isn't a role where leadership means stepping away from the technical work. You'll continue to design, code, validate and deliver solutions, including writing production-grade Python.
You'll also play a key role in evolving the data science environment towards Python-powered Jupyter workflows, using Claude Code and other LLM tools to accelerate research, development, testing, documentation, prototyping and production delivery.
The technology environment includes Python, PySpark, Delta Lake, Jupyter Notebooks, Azure Synapse, Trino, Power BI, Kubernetes-hosted model serving and FastAPI inference endpoints.
This is a technically ambitious, data-driven environment where Data Science is a core part of the product rather than a supporting function. The team values people who can combine technical depth with commercial judgement, communicate complex findings clearly and take ownership of outcomes.
It's particularly well suited to someone who enjoys staying hands-on while mentoring others and who is excited by the practical application of AI and LLM tooling within modern data science workflows.