Phyton with Spark Developer

Infotel UK Consulting

Chennai District

On-site

INR 600,000 - 900,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Infotel UK Consulting is looking for a Data Engineer based in Chennai, responsible for developing and optimizing data ingestion, transformation, and enrichment pipelines using Python, PySpark, and SQL. The role requires close collaboration with data architects and business analysts.

Key responsibilities include designing robust data pipelines, writing complex SQL queries, and maintaining CI/CD pipelines. Ideal candidates should have a Bachelor's degree along with strong technical skills in Python and SQL.

Qualifications

  • Strong knowledge about design patterns and development principles.
  • Hands-on experience with Python, including libraries like NumPy and pandas.
  • Experience with build tools and DevOps practices.

Responsibilities

  • Design and develop robust ingestion and transformation pipelines.
  • Write and optimize SQL queries for data aggregation.
  • Monitor production workloads and troubleshoot issues.

Skills

Python
PySpark
SQL
ETL

Education

Bachelor's degree or equivalent

Tools

Git
Jenkins
Docker
Maven
Bitbucket

Job description

Overview

Design and develop robust ingestion, transformation, and enrichment pipelines using Python, PySpark, and SQL. Collaborate with CEFS data architects, data scientists, and business analysts to translate functional requirements into technical specifications.

Responsibilities
  • Design & develop robust ingestion, transformation, and enrichment pipelines with Python, PySpark, and SQL
  • Write and optimize complex SQL queries, analytical UDFs, and window functions for data aggregation and reporting
  • Collaborate with CEFS data architects, data scientists, and business analysts to translate functional requirements into technical specifications
  • Unit-test, integrate-test, and review code
  • Maintain CI/CD pipelines (Git, Jenkins, Docker) for automated build, test, and deployment of jobs
  • Monitor production workloads and troubleshoot performance bottlenecks, memory issues, and job failures
  • Document data lineage, pipeline design, and operational run-books in Confluence/SharePoint
  • Keep up to date with latest technologies and trends and provide input, expertise and recommendations
Contributing Responsibilities
  • Contribute towards innovation (e.g. AI/ML); suggest new technical practices for efficiency improvement
  • Participate in Agile ceremonies (sprint planning, daily standups, retrospectives) and help groom the backlogs
  • Mentor junior engineers and champion best practices in Python coding, Spark optimization, and data-engineering patterns
  • Evaluate emerging technologies and deliver proof-of-concepts for CEFS
Technical & Behavioral Competencies
  • Resourceful to quickly understand complexities involved and provide the way forward
  • Experience in technical analysis of n-tier applications with multiple integrations using object oriented, APIs & Microservices approaches
  • Strong knowledge about design patterns and development principles
  • Inclination and prior experience of working across SQL, Python and ETL
  • Hands-on experience with Python (NumPy, pandas, Python Frameworks, Restful APIs, MS-SQL or Oracle)
  • PySpark - DataFrames, Spark SQL, Structured Streaming, performance tuning (partitioning, caching, broadcast joins)
  • Advanced SQL - complex queries, stored procedures, query optimization
  • Knowledge of Python packages such as Pandas, NumPy for data cleaning, wrangling, analysis and visualization
  • Experience in development and maintenance of code/scripts across applications, debugging, and production support
  • Knowledge of Linux/Unix environment (basic commands, shell scripting), testing, documentation and new framework
  • Experience with build tools like Maven and DevOps tools like Bitbucket, Jenkins
  • Knowledge of Agile, Scrum, DevOps
  • Development experience in a Data Engineering environment
  • Willingness to learn and work on diverse technologies
  • Self-motivated with good interpersonal skills and a drive to upgrade technologies
  • Good communication and coordination skills
Specific Qualifications
  • Good to have knowledge of front-end technologies, preferably Flask
Skills Referential
  • Technical Skills:
    • Python
    • PySpark
    • SQL
    • ETL
  • Behavioral Skills:
    • Ability to synthesize / simplify
    • Ability to collaborate / teamwork
    • Attention to detail / rigor
    • Ability to deliver / results driven

Education Level: Bachelor\'s degree or equivalent

Location: Chennai

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Phyton with Spark Developer
Phyton with Spark Developer

Infotel India • Chennai District

On-site
INR 800,000 - 1,200,000
Python with Spark Developer (5.1-7 years)-Chennai
Python with Spark Developer (5.1-7 years)-Chennai

Triwill Group • Chennai District

On-site
INR 1,200,000 - 2,400,000
Python with Spark Developer (5.1-7 years)-Chennai
Python with Spark Developer (5.1-7 years)-Chennai

Capco • Chennai District

On-site
INR 1,200,000 - 1,800,000
Data Engineer - SQL/PySpark
Data Engineer - SQL/PySpark

Forward Eye Technologies • Dadri

Hybrid
INR 1,400,000 - 2,000,000
Pyspark Developer
Pyspark Developer

Leading Global Technology Services Company • Chennai District, Bengaluru

Hybrid
INR 1,200,000 - 1,800,000
Data Engineer - ETL/PySpark
Data Engineer - ETL/PySpark

Forward Eye Technologies • Pune District

On-site
INR 1,200,000 - 2,400,000
PySpark Big Data Developer
PySpark Big Data Developer

Citi • Maharashtra

On-site
INR 1,100,000 - 1,800,000
Data Engineer Python - Senior Engineer
Data Engineer Python - Senior Engineer

Iris Software • Dadri

On-site
INR 1,800,000 - 2,400,000
Python/ETL Developer
Python/ETL Developer

Wissen • Bengaluru

Hybrid
INR 2,000,000 - 3,200,000
Python/ETL Developer
Python/ETL Developer

Wissen Technology • Bengaluru

Hybrid
INR 1,200,000 - 1,800,000