Data Analyst/Data Engineer Intern

DeHaat

Gurugram District

On-site

INR 167,000 - 279,000

Part time

43 hours ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

DeHaat is seeking a Data Analyst / Data Engineer Intern to join the Product & Technology team in Gurugram. You will build and maintain data pipelines, dashboards, reporting systems, and data-driven workflows in a hands-on role.

This internship focuses on databases, data pipelines, APIs, and data quality, with exposure to AWS RDS/Redshift and Frappe/ERPNext platforms. Duration is 6–9 months; immediate joining and work-from-office in Gurugram.

Qualifications

  • Strong knowledge of SQL, including joins, aggregations, CTEs and data manipulation.
  • Good working knowledge of Python for data processing and automation.
  • Experience with ETL/ELT pipelines and scheduled data jobs.
  • Familiarity with relational databases and data modelling.
  • Proficient in Excel/Google Sheets for data analysis.
  • Experience with REST APIs, JSON, and Postman.

Responsibilities

  • Build, maintain, and troubleshoot data pipelines across multiple sources (AWS RDS, Redshift, databases, APIs).
  • Write and optimize SQL queries for extraction, transformation, reconciliation, and reporting.
  • Develop Python scripts for ETL/ELT workflows and automation.
  • Manage and monitor scheduled cron jobs and recurring data workflows.
  • Identify data-quality issues and implement validation checks between sources and dashboards.
  • Work with APIs/JSON integrations to move data between apps.
  • Document data sources, schemas, transformations, and dependencies.
  • Create/maintain operational dashboards and internal tools using Frappe ERPNext.

Skills

SQL
Python
Data pipelines
APIs & JSON
Excel/Sheets
Problem solving
Data QA
GitHub
Frappe Framework

Tools

Postman
AWS RDS
Amazon Redshift
PostgreSQL
MySQL
GitHub
Linux
Cron jobs
ERPNext/Frappe

Job description

We are looking for a Data Analyst / Data Engineer Intern to work closely with the Product & Technology team on building and maintaining data pipelines, internal dashboards, reporting systems, and data-driven workflows.

This is a hands-on role suited for someone who is technically strong, comfortable working with databases and data pipelines, and interested in understanding how data is used to build operational products.

Key Responsibilities

A major part of the role will involve managing data flowing across multiple systems and databases.

  • Build, maintain, and troubleshoot data pipelines across sources such as **AWS RDS, Amazon Redshift, application databases, APIs, Google Sheets, and other internal systems**.
  • Write and optimize SQL queries for data extraction, transformation, reconciliation, and reporting.
  • Develop Python scripts for **ETL/ELT workflows, data transformation, validation, and automation**.
  • Manage and monitor scheduled cron jobs and recurring data workflows.
  • Identify pipeline failures, missing records, duplicates, schema inconsistencies, and other data-quality issues.
  • Create data validation and reconciliation checks between source systems and downstream dashboards.
  • Work with APIs and JSON-based integrations to move data between different applications.
  • Support documentation of data sources, schemas, transformations, and dependencies.

2. Dashboarding & Frappe-Based Applications

The second major responsibility will be building and maintaining operational dashboards and internal tools, primarily using Frappe Framework / ERPNext.

  • Develop and maintain dashboards, reports, DocTypes, forms, workflows, and views in Frappe.
  • Connect Frappe applications with databases, APIs, and internal data pipelines.
  • Build operational dashboards for Product, Technology, Field, and Business teams.
  • Translate business requirements into structured datasets, KPIs, filters, and dashboard components.
  • Debug data discrepancies between backend databases and dashboard outputs.
  • Improve dashboard usability, performance, and data freshness.
Additional Responsibilities
  • Use tools such as Postman for API testing and debugging.
  • Work with REST APIs, JSON payloads, authentication mechanisms, and webhook-based integrations.
  • Perform product and data QA before releases.
  • Support workflow automation across internal systems.
  • Use AI tools such as ChatGPT, Claude, or similar LLMs for coding assistance, data analysis, documentation, debugging, and workflow automation.
  • Prepare technical documentation, data dictionaries, SOPs, and implementation notes.
  • Quickly understand new systems and work across multiple technology platforms.
Preferred Skills
Must-have / Strongly Preferred
  • Strong knowledge of SQL, including joins, aggregations, CTEs, subqueries, and data manipulation.
  • Good working knowledge of Python, particularly for data processing and automation.
  • Understanding of relational databases and data modelling.
  • Experience working with ETL/ELT pipelines and scheduled data jobs.
  • Strong proficiency in Microsoft Excel / Google Sheets.
  • Understanding of REST APIs, JSON, and Postman.
  • Strong analytical and logical problem-solving skills.
Good to Have
  • Experience with Frappe Framework / ERPNext.
  • Exposure to AWS RDS, Amazon Redshift, PostgreSQL, MySQL, or similar databases.
  • Familiarity with cron jobs, schedulers, and basic Linux/server environments.
  • Experience building dashboards or internal reporting tools.
  • Familiarity with Git/GitHub.
  • Experience using AI coding/productivity tools such as ChatGPT, Claude, Cursor, or similar platforms.
Candidate Profile

We are looking for someone who:

  • Has strong computer science and data fundamentals.
  • Enjoys working with databases and solving data problems.
  • Can independently investigate why a pipeline, query, API, or dashboard is not working.
  • Pays close attention to data accuracy and edge cases.
  • Can work across engineering, product, and business requirements rather than operating only as a traditional analyst.
  • Is comfortable taking ownership of small technical projects from requirement gathering through implementation and testing.
  • Duration: 6-9 months
  • Location: Gurugram - Work from Office
  • Joining: Immediate
  • Team: Product & Technology

The role will provide significant hands-on exposure to production data pipelines, AWS databases, Frappe-based applications, APIs, operational dashboards, workflow automation, and AI-assisted development.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer Pune · Hybrid Engineering · Full-time →
Data Engineer Pune · Hybrid Engineering · Full-time →

Woodfrog Tech OPC Private Limited • Pune District

Hybrid
INR 1,800,000 - 2,400,000
Hybrid working
Data Engineer
Data Engineer

Access Corp. • Chennai District

On-site
INR 1,800,000 - 2,800,000
Data Analyst
Data Analyst

IIDE Education Private Limited • Mumbai

On-site
INR 600,000 - 900,000
Data Engineer
Data Engineer

KPG99 INC • Gurugram District

Hybrid
INR 1,200,000 - 1,800,000
Data Engineer
Data Engineer

Business Brio • Kolkata District

On-site
INR 900,000 - 1,800,000
Technical Lead - (Azure/AWS/GCP)
Technical Lead - (Azure/AWS/GCP)

Dentsu Global Services • Maharashtra

On-site
INR 1,200,000 - 2,000,000
Python Data Engineer
Python Data Engineer

ESR Healthcare • Bengaluru

On-site
INR 1,138,353 - 4,991,243
Senior Data Engineer
Senior Data Engineer

Calsoft • Pune District

On-site
INR 800,000 - 1,200,000
Data Engineer
Data Engineer

Weekday AI (YC W21) • Bengaluru

On-site
INR 700,000 - 1,500,000
Data Engineer
Data Engineer

Capri Global Capital (CGCL) • Gurugram District, Dadri

Hybrid
INR 900,000 - 1,300,000