Data Analyst/Data Engineer Intern

DeHaat

Gurugram District

On-site

INR 167,000 - 279,000

Full time

8 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

DeHaat is seeking a Data Analyst / Data Engineer Intern to join the Product & Technology team in Gurugram. You will build and maintain data pipelines, dashboards, and data-driven workflows using SQL, Python, and Frappe/ERPNext.

You will work on data integration across systems, API connections, and data quality checks, while contributing to documentation and workflow automation. Immediate joining for a 6–9 month internship is expected.

Qualifications

  • Strong knowledge of SQL including joins, aggregations, CTEs and data manipulation.
  • Good working knowledge of Python for data processing and automation.
  • Experience with ETL/ELT pipelines and scheduled data jobs.
  • Familiarity with REST APIs, JSON payloads, and Postman for testing.
  • Experience with Frappe Framework / ERPNext is a plus.

Responsibilities

  • Build, maintain, and troubleshoot data pipelines across multiple systems and databases.
  • Develop Python scripts for ETL/ELT workflows and automation.
  • Create dashboards and reports using Frappe/ERPNext, databases, and APIs.
  • Validate data quality and reconcile source systems with dashboards.
  • Collaborate with Product & Technology to document data sources and schemas.
  • Support basic frontend customization with JavaScript/HTML/CSS when needed.

Skills

SQL
Python
ETL/ELT pipelines
Data pipelines
REST APIs
Frappe/ERPNext
Excel/Sheets
JavaScript/HTML/CSS
AWS RDS/Redshift
Git/GitHub
Cron jobs

Tools

Postman
GitHub

Job description

Role Overview:

We are looking for a Data Analyst / Data Engineer Intern to work closely with the Product & Technology team on building and maintaining data pipelines, internal dashboards, reporting systems, and data-driven workflows.

This is a hands-on role suited for someone who is technically strong, comfortable working with databases and data pipelines, and interested in understanding how data is used to build operational products.

Key Responsibilities

1. Data Engineering & Data Pipelines

A major part of the role will involve managing data flowing across multiple systems and databases.

  • Build, maintain, and troubleshoot data pipelines across sources such as AWS RDS, Amazon Redshift, application databases, APIs, Google Sheets, and other internal systems.
  • Write and optimize SQL queries for data extraction, transformation, reconciliation, and reporting.
  • Develop Python scripts for ETL/ELT workflows, data transformation, validation, and automation.
  • Manage and monitor scheduled cron jobs and recurring data workflows.
  • Identify pipeline failures, missing records, duplicates, schema inconsistencies, and other data-quality issues.
  • Create data validation and reconciliation checks between source systems and downstream dashboards.
  • Work with APIs and JSON-based integrations to move data between different applications.
  • Support documentation of data sources, schemas, transformations, and dependencies.

2. Dashboarding & Frappe-Based Applications

The second major responsibility will be building and maintaining operational dashboards and internal tools, primarily using Frappe Framework / ERPNext.

  • Develop and maintain dashboards, reports, DocTypes, forms, workflows, and views in Frappe.
  • Connect Frappe applications with databases, APIs, and internal data pipelines.
  • Build operational dashboards for Product, Technology, Field, and Business teams.
  • Translate business requirements into structured datasets, KPIs, filters, and dashboard components.
  • Debug data discrepancies between backend databases and dashboard outputs.
  • Improve dashboard usability, performance, and data freshness.
  • Support basic frontend customisation using JavaScript, HTML, CSS, and Frappe client/server scripting where required.
Additional Responsibilities
  • Use tools such as Postman for API testing and debugging.
  • Work with REST APIs, JSON payloads, authentication mechanisms, and webhook-based integrations.
  • Perform product and data QA before releases.
  • Support workflow automation across internal systems.
  • Use AI tools such as ChatGPT, Claude, or similar LLMs for coding assistance, data analysis, documentation, debugging, and workflow automation.
  • Prepare technical documentation, data dictionaries, SOPs, and implementation notes.
  • Quickly understand new systems and work across multiple technology platforms.
##Preferred Skills
Must-have / Strongly Preferred
  • Strong knowledge of SQL, including joins, aggregations, CTEs, subqueries, and data manipulation.
  • Good working knowledge of Python, particularly for data processing and automation.
  • Understanding of relational databases and data modelling.
  • Experience working with ETL/ELT pipelines and scheduled data jobs.
  • Strong proficiency in Microsoft Excel / Google Sheets.
  • Understanding of REST APIs, JSON, and Postman.
  • Strong analytical and logical problem-solving skills.
Good to Have
  • Experience with Frappe Framework / ERPNext.
  • Exposure to AWS RDS, Amazon Redshift, PostgreSQL, MySQL, or similar databases.
  • Familiarity with cron jobs, schedulers, and basic Linux/server environments.
  • Knowledge of JavaScript, HTML, and CSS.
  • Experience building dashboards or internal reporting tools.
  • Familiarity with Git/GitHub.
  • Experience using AI coding/productivity tools such as ChatGPT, Claude, Cursor, or similar platforms.
## Candidate Profile

We are looking for someone who:

  • Has strong computer science and data fundamentals.
  • Enjoys working with databases and solving data problems.
  • Can independently investigate why a pipeline, query, API, or dashboard is not working.
  • Is comfortable learning unfamiliar technologies quickly.
  • Pays close attention to data accuracy and edge cases.
  • Can work across engineering, product, and business requirements rather than operating only as a traditional analyst.
  • Is comfortable taking ownership of small technical projects from requirement gathering through implementation and testing.
## Internship Details
  • Role: Data Analyst / Data Engineer Intern
  • Duration: 6–9 months
  • Location: Gurugram – Work from Office
  • Joining: Immediate
  • Team: Product & Technology

The role will provide significant hands-on exposure to production data pipelines, AWS databases, Frappe-based applications, APIs, operational dashboards, workflow automation, and AI-assisted development.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Digitraly • Chennai

On-site
Data Analyst - Intern
Data Analyst - Intern

Stirring Minds • New Delhi

On-site
INR 167,400 - 223,200
Data Analyst Intern at Enerture Technologies Pvt Ltd – Apply Now
Data Analyst Intern at Enerture Technologies Pvt Ltd – Apply Now

Enerture Technologies Pvt Ltd • Gurugram District

On-site
Intern - Business Transformation
Intern - Business Transformation

Cerebras • Bengaluru

On-site
INR 167,000 - 279,000
Mentor support
Exposure to production systems
Career path to full-time role
Data Engineer
Data Engineer

NeoStats • Bengaluru

On-site
Competitive Salary and Benefits
Ownership of Initiative
Continuous Coaching & Mentoring
+1
Intern- Python Developer, Newton
Intern- Python Developer, Newton

Affle • Dadri, Gurugram District

On-site
INR 1,200,000 - 1,800,000
Intern - Business Transformation
Intern - Business Transformation

Thoughtspot • Bengaluru

On-site
INR 167,000 - 446,000
Competitive internship compensation
Data Analyst - Fresher
Data Analyst - Fresher

Vyapar • Bengaluru

On-site
INR 300,000 - 400,000
Intern - Business Transformation
Intern - Business Transformation

ThoughtSpot • India

On-site
INR 201,000 - 335,000
Competitive internship compensation
Mentor support
Growth opportunities
Data Engineer 2
Data Engineer 2

WebSenor Ltd • Bengaluru

Hybrid
INR 900,000 - 1,400,000