Senior Data Consultant (Hadoop, Spark) (Kraków, PL, 30-302)

GFT Technologies SE

Polska

Hybrid

PLN 172,000 - 225,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

GFT Technologies SE is seeking a data engineer to join a technical team alongside Engineers, Data Analysts and Business Analysts, contributing to an Agile development process while designing, developing and maintaining scalable data solutions in a dynamic DevOps environment.

You will work on tasks including PySpark/Scala development, automating tests, code reviews, production support and building scalable data pipelines.

Qualifications

  • Experience with Pyspark or Scala development and design.
  • Experience using scheduling tools such as Airflow.
  • Knowledge of Hadoop ecosystem including Spark, Hive, YARN and ETL frameworks.
  • Strong SQL and RESTful services knowledge.
  • Experience working on Unix or Linux platforms.
  • Hands-on experience building data pipelines using Hadoop components.
  • Experience with Git, GitHub, Jenkins, Ansible and JIRA.
  • Understanding of big data modelling using relational and non-relational techniques.
  • Experience debugging code and communicating findings to development teams.
  • Openness to work 2 days a week from our client's office (Kraków).

Responsibilities

  • Define and contribute to software design and development using Pyspark.
  • Automate testing of new and existing components.
  • Promote development standards through code reviews and mentoring.
  • Provide production support and troubleshooting.
  • Implement tools and processes ensuring performance, scalability and monitoring.
  • Collaborate with Business Analysts to interpret and implement requirements.
  • Participate in planning, sprint reviews and retrospectives.
  • Contribute to system architecture and design.

Skills

Pyspark/Scala development
Airflow
Hadoop ecosystem
SQL
Unix/Linux
Data pipelines
Git/GitHub/Jenkins/Ansible/JIRA
Big data modeling
Debugging code
Kraków on-site (2 days)

Tools

Airflow
Hadoop
Spark
Hive
YARN
ETL frameworks
Git
GitHub
Jenkins
Ansible
JIRA

Job description

Type of contract: B2B contract
Salary range: 125-163 PLN net/h


What will you do?

You will work as a key member of a technical team alongside Engineers, Data Analysts and Business Analysts, contributing to a collaborative Agile development process while designing, developing and maintaining scalable data solutions in a dynamic DevOps environment.


Your tasks


  • Define and contribute to software design and development using Pyspark

  • Automate testing of new and existing components

  • Promote development standards through code reviews and mentoring

  • Provide production support and troubleshooting

  • Implement tools and processes ensuring performance, scalability and monitoring

  • Collaborate with Business Analysts to interpret and implement requirements

  • Participate in planning, sprint reviews and retrospectives

  • Contribute to system architecture and design


Requirements


  • Experience with Pyspark or Scala development and design

  • Experience using scheduling tools such as Airflow

  • Knowledge of Hadoop ecosystem including Spark, Hive, YARN and ETL frameworks

  • Strong SQL and RESTful services knowledge

  • Experience working on Unix or Linux platforms

  • Hands-on experience building data pipelines using Hadoop components

  • Experience with Git, GitHub, Jenkins, Ansible and JIRA

  • Understanding of big data modelling using relational and non-relational techniques

  • Experience debugging code and communicating findings to development teams

  • Openness to work 2 days a week from our client's office (Kraków)


Nice to have


  • Experience with Elasticsearch

  • Experience developing Java APIs

  • Experience in data ingestion processes

  • Understanding of cloud design patterns

  • Exposure to DevOps and Agile methodologies such as Scrum and Kanban

  • Experience with Spark streaming

  • Experience with Apache Airflow in production

  • Experience with Hadoop ecosystem in enterprise environments

  • Knowledge of Python backend services

  • Experience with Scala for high performance systems

  • Experience in data integration and ETL processes

  • Knowledge of PL/SQL

  • Experience with Linux and Unix system operations


Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Consultant (Hadoop, Spark)
Senior Data Consultant (Hadoop, Spark)

GFT TECHNOLOGIES SE • Kraków

On-site
PLN 172,200 - 224,548
Senior Data Engineer (Hadoop, Spark) (Kraków, PL, 30-302)
Senior Data Engineer (Hadoop, Spark) (Kraków, PL, 30-302)

GFT Technologies SE • Polska

Hybrid
PLN 176,000 - 271,000
Hybrid work Krakow
Medical benefits
Lunch subsidy
+4
Expert Java/Scala Consultant with Spark (Kraków, PL, 30-302)
Expert Java/Scala Consultant with Spark (Kraków, PL, 30-302)

GFT Technologies SE • Polska

Hybrid
PLN 214,000 - 282,000
Lead Data Consultant (Kraków, PL, 30-302)
Lead Data Consultant (Kraków, PL, 30-302)

GFT Technologies SE • Kraków

Hybrid
PLN 214,000 - 287,000
Expert Data Consultant (Kraków, PL, 30-302)
Expert Data Consultant (Kraków, PL, 30-302)

GFT Technologies SE • Kraków

Hybrid
PLN 214,000 - 282,000
Senior Java ETL Consultant (Kraków, PL, 30-302)
Senior Java ETL Consultant (Kraków, PL, 30-302)

GFT Technologies SE • Kraków

Hybrid
PLN 172,000 - 220,000
Hybrid work arrangement
Tech Lead Java Engineer with Spark (Kraków, PL, 30-302)
Tech Lead Java Engineer with Spark (Kraków, PL, 30-302)

GFT Technologies SE • Kraków

Hybrid
PLN 17,000 - 27,000
Hybrid schedule
Medical insurance
Lunch subsidy
+2
Senior Data Consultant
Senior Data Consultant

GFT TECHNOLOGIES SE • Łódź

Hybrid
Confidential
Big Data Engineer (Spark/Scala) Spark, Scala, Python, Hadoop, SQL, AWS Warsaw, Łódź, Gdańsk, Gdynia
Big Data Engineer (Spark/Scala) Spark, Scala, Python, Hadoop, SQL, AWS Warsaw, Łódź, Gdańsk, Gdynia

Diverse CG Sp. z o.o. Sp.k. • Łódź, Gdynia

On-site
PLN 80,000 - 100,000
Private medical care
Co-financing for the sports card
Constant support of dedicated consultant
Big Data Engineer (Spark/Scala) Spark, Scala, Python, Hadoop, SQL, AWS Warsaw, Łódź, Gdańsk, Gdynia
Big Data Engineer (Spark/Scala) Spark, Scala, Python, Hadoop, SQL, AWS Warsaw, Łódź, Gdańsk, Gdynia

DCG Poland • Łódź, Gdynia

On-site
PLN 167,400 - 279,000
Private medical care
Co-financing for the sports card
Constant support of dedicated consultant