Data Engineer — Recommendation Engine

Remotestar

Gurugram District

On-site

INR 800,000 - 1,500,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Competitive salary
Equity opportunities

Job summary

Remotestar is looking for a Data Engineer in Gurgaon to build and maintain data pipelines for their recommendation engine. You'll be responsible for external data sources, developing the Deal Quality Score pipeline, and ensuring data quality for the engine’s operations.

Ideal candidates have 3–5 years of experience, strong Python skills, and familiarity with AWS services. This foundational role offers competitive compensation and equity opportunities.

Qualifications

  • 3–5 years of data engineering experience, ideally at a startup or product company.
  • Experience building and maintaining web scrapers or data collection pipelines at scale.
  • Hands-on experience with a workflow orchestrator like Airflow or Prefect.
  • Solid SQL skills; experience with ClickHouse or other OLAP databases is a plus.
  • Familiarity with Redis as a serving layer.

Responsibilities

  • Own external data pipelines for price benchmarking and trend signals.
  • Build the Deal Quality Score pipeline for product competitiveness scoring.
  • Design and own the in-store event schema in ClickHouse.
  • Maintain data quality and freshness SLAs.

Skills

Python
SQL
Web scraping
Data pipeline building
Airflow
Redis
AWS (S3, Glue, ECS)

Tools

ClickHouse
Spark

Job description

The Role

We're building a recommendation engine that surfaces the right products to the right player at the right moment inside our in-game store. A core challenge: we have limited in-app behavioural data today, so the system must rely heavily on external market signals to make great recommendations from day one.

As our first Data Engineer, you will own the data infrastructure that makes this possible. You'll build the pipelines that collect, process, and serve these signals into our real-time ranking system.

This is a foundational hire. The entire recommendation engine—from Deal Quality Scores to seasonal trends—depends on the pipelines you build.

What You'll Do
  • Own external data pipelines—scrapers for Flipkart/Amazon bestseller rankings, PriceHunt/Smartprix for price benchmarking, Google Trends API for brand and category demand signals
  • Build the Deal Quality Score pipeline—a daily batch job that computes a competitiveness score for every product in our catalogue, stored in Redis for sub-millisecond lookup at serving time
  • Maintain a seasonal and festive calendar—structured data store for trend overlays (IPL, Diwali, back-to-school, etc.)
  • Design and own the in-store event schema in ClickHouse that will power behavioural cohorts as in-app data
  • Build ETL infrastructure (S3 + Spark/Glue) for longer-horizon trend and market data
  • Own data quality and freshness SLAs—you are responsible when a signal the reco engine depends on breaks silently
What We're Looking For
  • 3–5 years of data engineering experience, ideally at a startup or product company
  • Strong Python—you write clean, production‑grade pipeline code, not just notebooks
  • Experience building and maintaining web scrapers or data collection pipelines at scale
  • Hands‑on experience with a workflow orchestrator—Airflow, Prefect, or equivalent
  • Solid SQL; experience with ClickHouse or another OLAP database is a strong plus
  • Familiarity with Redis as a serving layer—you understand TTL, key design, and cache invalidation
  • Comfortable with AWS—S3, Glue, ECS; you can set up infra without needing DevOps help
  • You care about data quality—you monitor pipelines, set up alerts, and feel responsible when something breaks
Strong Plus (Nice to Have)
  • Experience with e-commerce or marketplace data—price intelligence, product catalogues, category taxonomy
  • Familiarity with recommender system data patterns
  • Experience with Spark or distributed processing for larger datasets
  • Prior work on gaming or consumer mobile products

Location: Gurgaon

Compensation: Competitive salary with an opportunity to get meaningful equity

Reports to: CTO

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Applied ML Engineer
Applied ML Engineer

Remotestar • Gurugram District

On-site
INR 1,500,000 - 2,000,000
Senior Data Engineer
Senior Data Engineer

BookMyMentor • Gurgaon

On-site
INR 1,200,000 - 2,000,000
Principal Data Engineer
Principal Data Engineer

HG Insights • Pune District

On-site
INR 4,000,000 - 6,400,000
Data Engineer
Data Engineer

Keka Inc. • Karnataka

On-site
INR 1,200,000 - 2,400,000
Data Engineer
Data Engineer

e-Stone Information Technology Private Limited • Mumbai

On-site
INR 1,500,000 - 2,800,000
Data Engineer - Python / Big Data
Data Engineer - Python / Big Data

Darwinbox Digital Solutions Pvt. Ltd. • Bengaluru

On-site
INR 1,800,000 - 2,400,000
Senior Software Engineer II - Data Engineering & Platform
Senior Software Engineer II - Data Engineering & Platform

Playsimple Games Private Limited • Bengaluru

On-site
INR 4,000,000 - 7,000,000
Data Engineer
Data Engineer

dentsu • New Delhi

On-site
INR 900,000 - 1,500,000
Data Engineer
Data Engineer

xtsworld • Pune District

On-site
INR 1,500,000 - 2,000,000
Free meals
Collaborative work culture
Health benefits
Senior Software Engineer II - Data Engineering & Platform
Senior Software Engineer II - Data Engineering & Platform

Zoho • Bengaluru Urban

Hybrid
INR 3,500,000 - 6,000,000