Big Data Engineer

Qcentrio

Dadri

On-site

INR 1,500,000 - 2,100,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Qcentrio is seeking an experienced Data Engineer to join our team in India. You will design and implement scalable data pipelines, improve data quality, and collaborate with stakeholders across development and on-site teams.

You should have 5+ years of experience building data lake solutions on AWS (S3, EMR, Hive, PySpark), strong SQL and Python skills, and familiarity with Airflow and GitHub. A Bachelor’s degree in Computer Science is required. Immediate availability.

Qualifications

  • 5+ years of data engineering experience.
  • Experience with AWS, EMR, S3, Hive & PySpark.
  • Proficiency in SQL and Python.

Responsibilities

  • Actively participate in all phases of the software development lifecycle.
  • Solve complex business problems by applying disciplined development methods.
  • Produce scalable, flexible, and maintainable data solutions.
  • Analyse source and target data; map transformations to meet requirements.
  • Interact with clients and onsite coordinators during project phases.
  • Design and implement features with stakeholders in business and technology.
  • Improve data quality by anticipating and solving data management issues.
  • Clean, prepare, and optimize data at scale for ingestion and use.
  • Support new data management projects and restructure existing data architecture.
  • Implement automated workflows using scheduling tools.
  • Follow CI, TDD, and production deployment frameworks.
  • Contribute to design, code, and test plans for data pipelines.
  • Analyze and profile data for scalable solutions.
  • Troubleshoot data issues and perform root-cause analysis.

Skills

Big Data
Python
SQL
Spark/PySpark
AWS Cloud

Education

Bachelor's degree in computer science

Tools

GitHub
Airflow
Databricks
Hive
PySpark

Job description

Notice Period : Immediate - 30 days

Mandatory Skills : Big Data, Python, SQL, Spark/Pyspark, AWS Cloud

JD and required Skills & Responsibilities :

  • - Actively participate in all phases of the software development lifecycle, including requirements gathering, functional and technical design, development, testing, roll-out, and support.
  • - Solve complex business problems by utilizing a disciplined development methodology.
  • - Produce scalable, flexible, efficient, and supportable solutions using appropriate technologies.
  • - Analyse the source and target system data. Map the transformation that meets the requirements.
  • - Interact with the client and onsite coordinators during different phases of a project.
  • - Design and implement product features in collaboration with business and Technology stakeholders.
  • - Anticipate, identify, and solve issues concerning data management to improve data quality.
  • - Clean, prepare, and optimize data at scale for ingestion and consumption.
  • - Support the implementation of new data management projects and re-structure the current data architecture.
  • - Implement automated workflows and routines using workflow scheduling tools.
  • - Understand and use continuous integration, test-driven development, and production deployment frameworks.
  • - Participate in design, code, test plans, and dataset implementation performed by other data engineers in support of maintaining data engineering standards.
  • - Analyze and profile data for the purpose of designing scalable solutions.
  • - Troubleshoot straightforward data issues and perform root cause analysis to proactively resolve product issues.

Required Skills :

  • - 5+ years of relevant experience developing Data and analytic solutions.
  • - Experience building data lake solutions leveraging one or more of the following AWS, EMR, S3, Hive & PySpark
  • - Experience with relational SQL.
  • - Experience with scripting languages such as Python.
  • - Experience with source control tools such as GitHub and related dev process.
  • - Experience with workflow scheduling tools such as Airflow.
  • - In-depth knowledge of AWS Cloud (S3, EMR, Databricks)
  • - Has a passion for data solutions.
  • - Has a strong problem-solving and analytical mindset
  • - Working experience in the design, Development, and test of data pipelines.
  • - Experience working with Agile Teams.
  • - Able to influence and communicate effectively, both verbally and in writing, with team members and business stakeholders
  • - Able to quickly pick up new programming languages, technologies, and frameworks.
  • - Bachelor's degree in computer science
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Big Data Engineer
Big Data Engineer

Qcentrio • Gurugram District

On-site
INR 2,400,000 - 4,200,000
Big Data Engineer - Python / PySpark
Big Data Engineer - Python / PySpark

Qcentrio • Bengaluru

On-site
INR 1,200,000 - 2,400,000
Big Data Engineer - Python/PySpark
Big Data Engineer - Python/PySpark

Qcentrio • Jaipur

On-site
INR 1,500,000 - 2,100,000
Big Data Engineer - Python/PySpark
Big Data Engineer - Python/PySpark

Qcentrio • Surat

On-site
INR 1,200,000 - 2,400,000
Big Data Engineer - Python/PySpark
Big Data Engineer - Python/PySpark

Qcentrio • Dadri

On-site
INR 1,400,000 - 2,200,000
Big Data Engineer
Big Data Engineer

Qcentrio • Jaipur

On-site
INR 1,500,000 - 2,100,000
Big Data Engineer
Big Data Engineer

Qcentrio • Pune District

Hybrid
INR 2,500,000 - 3,600,000
Big Data Engineer
Big Data Engineer

Qcentrio • Bengaluru

On-site
INR 1,500,000 - 2,600,000
Big Data Engineer - Python/PySpark
Big Data Engineer - Python/PySpark

Qcentrio • Mumbai

On-site
INR 2,500,000 - 4,000,000
AWS Data Engineer
AWS Data Engineer

Zorba AI • Chennai District

On-site
INR 800,000 - 1,800,000