[C3F] Data Platform Engineer

Visa Hunt

Poland

On-site

PLN 180,000 - 240,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Flexible work
International projects
Private healthcare
Multisport card
Language classes

Job summary

Software Mind develops solutions that make an impact for companies around the globe. We seek a Data/Platform Engineer to build and operate CDC pipelines, mainly from PostgreSQL into Azure, using Kafka Connect and Debezium.

You’ll tune Kubernetes workloads, automate onboarding with Python, and integrate with Event Hubs, ADLS, ADF, and Databricks. Ideal candidates have hands-on streaming experience, strong SQL/PostgreSQL skills, and proficiency with Azure data services.

Qualifications

  • Solid commercial experience as a data or platform engineer, with hands-on work on streaming or CDC pipelines rather than batch reporting alone.
  • Practical Kafka knowledge - topics, partitions, offsets, consumer groups and delivery semantics - including at least one Kafka Connect deployment you ran yourself
  • Strong SQL and PostgreSQL skills, with working knowledge of WAL, logical replication, replication slots and replication lag
  • Working understanding of CDC concepts: initial snapshots, inserts, updates and deletes, event ordering, at-least-once delivery, and schema evolution
  • Confident Python for automation and tooling - orchestration, monitoring, recovery scripts, and automated tests
  • Hands-on experience with Azure data services, for example Event Hubs, ADLS or Azure PostgreSQL
  • Comfortable working with Kubernetes as a user: deploying workloads, handling configuration and secrets, reading logs, debugging failing pods
  • Ability to debug a running pipeline from metrics and logs - telling throughput problems from backpressure, retries or a genuine connector failure

Responsibilities

  • Build and operate change-data-capture pipelines from PostgreSQL into Azure, using Kafka Connect and Debezium as the core of the platform
  • Configure, deploy and scale connectors end to end - connector setup, task management, offsets, schema history, and snapshot strategy
  • Run these pipelines as stateful workloads on Kubernetes (AKS), covering configuration, secrets, networking and resource tuning
  • Monitor and troubleshoot the platform in production: connector failures, task rebalances, restarts, throughput and backpressure, message-size limits, retries and recovery
  • Automate the platform in Python - configuration-driven onboarding of new data sources, pipeline orchestration, monitoring and alerting, recovery workflows, and automated testing
  • Integrate CDC streams with the wider Azure data stack: Event Hubs, ADLS, Azure PostgreSQL, ADF and Databricks
  • Manage platform infrastructure as code, so environments are reproducible and changes are reviewable
  • Apply data protection requirements to sensitive data flowing through the pipelines - masking, hashing, access control and retention

Skills

Kafka
PostgreSQL
Python
Azure
Kubernetes
Data pipelines
CDC concepts

Tools

Kafka Connect
Debezium
Terraform
Bicep
Databricks
ADF

Job description

Project – the aim you’ll have

Our customer provides innovative solutions and insights that enable our clients to manage risk and hire the best talent. Their advanced global technology platform supports fully scalable, configurable screening programs that meet the unique needs of over 33,000 clients worldwide. Headquartered in Atlanta, GA, they have an internationally distributed workforce spanning 19 countries with about 5,500 employees. Our partner perform over 93 million screens annually in over 200 countries and territories.

Position – how you’ll contribute
  • Build and operate change-data-capture pipelines from PostgreSQL into Azure, using Kafka Connect and Debezium as the core of the platform
  • Configure, deploy and scale connectors end to end - connector setup, task management, offsets, schema history, and snapshot strategy
  • Run these pipelines as stateful workloads on Kubernetes (AKS), covering configuration, secrets, networking and resource tuning
  • Monitor and troubleshoot the platform in production: connector failures, task rebalances, restarts, throughput and backpressure, message-size limits, retries and recovery
  • Automate the platform in Python - configuration-driven onboarding of new data sources, pipeline orchestration, monitoring and alerting, recovery workflows, and automated testing
  • Integrate CDC streams with the wider Azure data stack: Event Hubs, ADLS, Azure PostgreSQL, ADF and Databricks
  • Manage platform infrastructure as code, so environments are reproducible and changes are reviewable
  • Apply data protection requirements to sensitive data flowing through the pipelines - masking, hashing, access control and retention
Expectations – the experience you need
  • Solid commercial experience as a data or platform engineer, with hands-on work on streaming or CDC pipelines rather than batch reporting alone
  • Practical Kafka knowledge - topics, partitions, offsets, consumer groups and delivery semantics - including at least one Kafka Connect deployment you ran yourself
  • Strong SQL and PostgreSQL skills, with working knowledge of WAL, logical replication, replication slots and replication lag
  • Working understanding of CDC concepts: initial snapshots, inserts, updates and deletes, event ordering, at-least-once delivery, and schema evolution
  • Confident Python for automation and tooling - orchestration, monitoring, recovery scripts, and automated tests
  • Hands-on experience with Azure data services, for example Event Hubs, ADLS or Azure PostgreSQL
  • Comfortable working with Kubernetes as a user: deploying workloads, handling configuration and secrets, reading logs, debugging failing pods
  • Ability to debug a running pipeline from metrics and logs - telling throughput problems from backpressure, retries or a genuine connector failure
Additional skills – the edge you have
  • Production experience with Debezium specifically - snapshot strategies on large tables, schema history recovery, offset loss, and bringing connectors back after failure
  • Experience operating stateful workloads on AKS: StatefulSets, stable worker identity, and resource tuning under load
  • Infrastructure-as-code and CI/CD for data platform components (Terraform, Bicep or similar)
  • Hands-on work with Databricks and ADF at production scale
  • Experience implementing data protection controls for sensitive data - masking, hashing, access control and retention policies
Our offer – professional development, personal growth
  • Flexible employment and remote work
  • International projects with leading global clients
  • International business trips
  • Non-corporate atmosphere
  • Language classes
  • Internal & external training
  • Private healthcare and insurance
  • Multisport card
  • Well-being initiatives

Software Mind develops solutions that make an impact for companies around the globe. Tech giants & unicorns, transformative projects, emerging technologies and limitless opportunities – these are a few words that describe an average day for us. Building cross-functional engineering teams that take ownership and crave more means we’re always on the lookout for talented people who bring passion and creativity to every project. Our culture embraces openness, acts with respect, shows grit & guts and combines employment with enjoyment.

Originally posted on Himalayas

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Remote Data Platform Engineer - Azure, Kafka & CDC
Remote Data Platform Engineer - Azure, Kafka & CDC

Visa Hunt • Poland

On-site
PLN 180,000 - 240,000
Flexible work
International projects
Private healthcare
+2
Data Engineer / Senior Data Engineer with Azure, Databricks and Microsoft Fabric
Data Engineer / Senior Data Engineer with Azure, Databricks and Microsoft Fabric

DataArt • Poland

Hybrid
PLN 200,000 - 320,000
Vacation days 26+
Health and life insurance (Luxmed)
MyBenefit with Multisport
+9
Data Engineer (Databricks)
Data Engineer (Databricks)

Addepto • Poland

Remote
PLN 80,000 - 100,000
Flexible work arrangements
20 fully paid days off
Medical and sports packages
+1
Lead Azure Data Engineer with Databricks
Lead Azure Data Engineer with Databricks

Lingaro • Poland

Hybrid
PLN 254,000 - 340,000
Stable employment
"Office as an option" model
Workation opportunities
+2
Senior Data Engineer (Azure/Databricks)
Senior Data Engineer (Azure/Databricks)

HeadHR • Poznań

On-site
PLN 120,000 - 150,000
Private medical care
Co-financing for the sports card
Constant support of dedicated consultant
+1
Senior Data Quality Engineer, SDET with Databricks and Microsoft Fabric
Senior Data Quality Engineer, SDET with Databricks and Microsoft Fabric

DataArt • Poland

Hybrid
PLN 120,000 - 190,000
Vacation days: Up to 26 business days
Health and life insurance (Luxmed)
English language classes from day one
+5
Senior Data Engineer - Real Time Streams
Senior Data Engineer - Real Time Streams

Globaldev Group • Poland

On-site
PLN 230,000 - 320,000
20 days paid vacation
5 sick days
Flexible schedule
+3
Senior Data Engineer
Senior Data Engineer

DEVTALENTS Sp. z o.o. • Województwo mazowieckie

On-site
PLN 150,000 - 200,000
Training opportunities
Supportive culture
Python Data Platform Engineer
Python Data Platform Engineer

Cisco • Kraków

On-site
PLN 180,000 - 260,000
Senior Data Engineer ID86297
Senior Data Engineer ID86297

AgileEngine • Warszawa

On-site
PLN 260,000 - 360,000
Growth without limits
Competitive compensation
Remote work 100% with flexible hours
+3