Data Engineer – AI, Java, Python, Spark

Jobtailor

Warszawa

On-site

PLN 180,000 - 280,000

Full time

13 days ago

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Jobtailor in Warsaw seeks a Senior Data Engineer to design, build and optimize scalable data pipelines and cloud-native services. You will work with Python/Java, Spark, Databricks and Snowflake to deliver production-grade solutions in a fast-paced environment.

You’ll lead cross-functional initiatives, apply GenAI/LMM techniques, and ensure reliability through testing and observability, collaborating with data scientists, software engineers, and product teams.

Qualifications

  • 5+ years of professional software/data engineering experience.
  • Strong hands-on experience with Python and Java.
  • Strong experience with Apache Spark and distributed data processing.
  • Experience with Databricks and/or modern lakehouse platforms.
  • Experience with Snowflake or comparable cloud data warehouses.
  • Practical experience with Kubernetes and cloud-native technologies.
  • Experience designing and maintaining large-scale data pipelines.
  • Strong understanding of distributed systems, scalability and production engineering.
  • Experience developing ML/AI or GenAI applications.
  • Experience with LLM-based applications, RAG, AI agents or LLM orchestration.
  • Familiarity with LangChain, LangGraph or similar GenAI frameworks.
  • Strong software engineering fundamentals including testing, code quality and system design.

Responsibilities

  • Develop, test and maintain high-quality, production-ready software
  • Design and implement large-scale data pipelines and distributed processing systems
  • Build scalable cloud-native services and platforms
  • Provide technical leadership for cross-team initiatives and complex engineering projects
  • Design and develop reusable libraries, frameworks and platform components
  • Optimize distributed data processing workloads for performance, scalability and reliability
  • Work with Databricks, Apache Spark and Snowflake data platforms
  • Develop and deploy applications using Python and/or Java
  • Build and operate containerized workloads using Kubernetes and cloud-native technologies
  • Design and implement GenAI/LLM-based applications and services
  • Use LangChain and LangGraph for LLM orchestration and agentic workflows
  • Collaborate with data scientists, software engineers, architects and product teams
  • Establish engineering best practices around testing, observability, reliability and deployment

Skills

Python
Java
Apache Spark
Kubernetes
GenAI
LangChain
LangGraph

Tools

Databricks
Snowflake
LangChain
LangGraph

Job description

  • Develop, test and maintain high-quality, production-ready software
  • Design and implement large-scale data pipelines and distributed processing systems
  • Build scalable cloud-native services and platforms
  • Provide technical leadership for cross-team initiatives and complex engineering projects
  • Design and develop reusable libraries, frameworks and platform components
  • Optimize distributed data processing workloads for performance, scalability and reliability
  • Work with Databricks, Apache Spark and Snowflake data platforms
  • Develop and deploy applications using Python and/or Java
  • Build and operate containerized workloads using Kubernetes and cloud-native technologies
  • Design and implement GenAI/LLM-based applications and services
  • Use LangChain and LangGraph for LLM orchestration and agentic workflows
  • Collaborate with data scientists, software engineers, architects and product teams
  • Establish engineering best practices around testing, observability, reliability and deployment
Requirements
  • 5+ years of professional software/data engineering experience
  • Strong hands-on experience with Python and/or Java
  • Strong experience with Apache Spark and distributed data processing
  • Experience with Databricks and/or modern lakehouse platforms
  • Experience with Snowflake or comparable cloud data warehouses
  • Practical experience with Kubernetes and cloud-native technologies
  • Experience designing and maintaining large-scale data pipelines
  • Strong understanding of distributed systems, scalability and production engineering
  • Experience developing ML/AI or GenAI applications
  • Experience with LLM-based applications, RAG, AI agents or LLM orchestration
  • Familiarity with LangChain, LangGraph or similar GenAI frameworks
  • Strong software engineering fundamentals including testing, code quality and system design
Core Competencies

Demonstrates expertise in developing and maintaining high-quality software, with a strong focus on building scalable cloud-native services and optimizing distributed data processing. Proficient in Python, Java, and modern data platforms, with a solid understanding of GenAI applications and engineering best practices.

Highest-signal resume keywords
  • Python Development
  • Java Development
  • Apache Spark
  • Kubernetes
  • GenAI Applications
ATS Optimization Keywords
Hard Skills
  • Software Engineering
  • Data Engineering
  • Distributed Systems
  • Data Pipeline Design
  • Performance Optimization
  • Testing and Code Quality
  • System Design
  • ML/AI Development
  • LLM Orchestration
  • Cloud-Native Technologies
Soft Skills
  • Technical Leadership
  • Collaboration
Industry Keywords
  • Cloud Data Warehouses
  • Lakehouse Platforms
  • Distributed Processing Systems
  • Engineering Best Practices
Tools & Technologies
  • Databricks
  • Snowflake
  • LangChain
  • LangGraph
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Lead - Client Technology
Data Lead - Client Technology

Ernst & Young Advisory Services Sdn Bhd • Wrocław

On-site
PLN 276,000 - 362,000
Senior Data Engineer
Senior Data Engineer

Sigmasoftware2 • Poland

On-site
PLN 180,000 - 280,000
Senior Data Engineer: Cloud, GenAI & Scalable Pipelines
Senior Data Engineer: Cloud, GenAI & Scalable Pipelines

eTeam • Warszawa

On-site
PLN 256,000 - 300,000
Python/AI Engineer
Python/AI Engineer

Jobtailor • Warszawa

On-site
PLN 180,000 - 240,000
Principal Gen AI Software Engineer
Principal Gen AI Software Engineer

Jobtailor • Kraków

On-site
PLN 300,000 - 600,000
Data Scientist Lead
Data Scientist Lead

Jobtailor • Województwo pomorskie

On-site
PLN 300,000 - 540,000
AI Data Engineer (Data Engineering, Cloud Platform, Python)
AI Data Engineer (Data Engineering, Cloud Platform, Python)

Capgemini • Warszawa

Hybrid
PLN 120,000 - 160,000
Medical care
Insurance
Wellness resources
+1
Data Engineer
Data Engineer

Avanade • Kraków

On-site
PLN 60,000 - 80,000
Senior Software Engineer – Technical Lead
Senior Software Engineer – Technical Lead

Jobtailor • Kraków

On-site
PLN 180,000 - 240,000
Senior Data Engineer (Databricks Migration)
Senior Data Engineer (Databricks Migration)

Sigma Software • Kraków

On-site
PLN 240,000 - 360,000