Data Engineer

Recro

Bengaluru

On-site

INR 2,500,000 - 4,200,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Recro in Bengaluru is seeking a Data Engineer II (SDE-2) to join our data team. The role focuses on building a high-performance Lakehouse, managing CDC flows, and optimizing distributed query engines. You will work with engineering and SRE to keep production systems stable and scalable.

This engineering-heavy position requires strong Python/PySpark, SQL, and experience with PySpark, Flink, and AWS services. You'll design domain models for OLAP and lead technical initiatives.

Qualifications

  • Experience: 3–5 years in Data Engineering, distributed systems and cloud-native architectures.
  • Coding: Expert-level Python/PySpark and SQL.
  • Familiarity with Go/Java/Scala is a plus.
  • Infrastructure: Hands-on experience with AWS (S3, EKS, MSK) and IaC.
  • Orchestration: Experience with Airflow or Temporal for complex workflows.
  • AI-Native: Proficiency in using AI tools (Claude, Codex, Copilot) to write, test, and document code.
  • Systems Thinking: Ability to explain trade-offs between storage formats and processing frameworks.
  • Tech Leadership: Drive key tech initiatives by preparing TRD and design reviews.
  • Domain Modelling: Hands-on design of OLAP domain models (Fact, Dimension, SCDs, OBT patterns).
  • Self Starter: Lead the team technically and bring in new ideas to contribute to growth.
  • Customer First: Interact with Product & Key Stakeholders to add value to business workflow with data & analytics.

Responsibilities

  • Handle first-level technical investigation and mitigation for production issues by executing predefined command playbooks.
  • Collaborate with engineering/SRE teams to keep customer-facing systems stable and healthy.

Skills

Python
PySpark
SQL
Go
Java
Scala
Airflow
Temporal
AWS
EMR
EKS
MSK
S3
Terraform
Trino
ClickHouse
Flink
CDC

Tools

GitLab
GitHub Actions
Terraform
Kubernetes

Job description

eThe Product Support Engineer will handle first-level technical investigation and mitigation for production issues by executing predefined command playbooks. You will work closely with engineering/SRE teams to keep customer-facing systems stable and healthy. We are looking for a Data Engineer II (SDE-2) to join our data team. The ideal candidate will be a play a key role to develop of high performant and scalable Data Lake-house, moving us toward a world of sub-minute data latency and unified batch/streaming compute. This is an engineering-heavy role where you will manage complex CDC flows, optimize distributed query engines and leverage AI to accelerate our development lifecycle

.Technical Prioritie
  • sReal-time CDC: Ownership of high-throughput ingestion from RDBMS to Lakehouse using Debezium, PeerDB
  • .Lakehouse Architecture: Designing and optimizing table formats (Iceberg, Delta, Hudi) for both performance and storage efficiency
  • .Unified Compute: Developing robust ETL/ELT frameworks in PySpark and Flink (handling both batch and streaming workloads)
  • .Infrastructure & Ops: Managing data workloads on AWS (EMR, EKS, MSK, S3) and automating everything via Gitlab/Github Actions
  • .Query & BI: Tuning Trino or Clickhouse to power real-time dashboards in Metabase, Superset, and PowerBI
.Requirement
  • sExperience: 3–5 years in Data Engineering, specifically with distributed systems and cloud-native architectures
  • .Coding: Expert-level Python/PySpark and SQL
  • .Familiarity with Go/Java/Scala is a plu
  • sInfrastructure: Hands‑on experience with AWS (S3, EKS, MSK) and Infrastructure-as-Code
  • .Orchestration: Experience with Airflow or Temporal for complex workflow management
  • .AI-Native: Proficiency in using AI tools (Claude, Codex, Copilot) to write, test, and document code efficiently
  • .Systems Thinking: Ability to explain the trade-offs between different storage formats and processing frameworks
  • .Tech LeaderShip : Drive key tech initiatives by preparing TRD and actively involve in design reviews
  • .Domain Modelling - Should be hands on in designing Domain models for OLAP like Fact, Dimension and types of SCD’s and OBT pattern tables
  • .Self Starter - Lead the team technically and bring in new ideas to contribute to the growth of the charter
  • .Customer First - Interact with the Product & Key Stakeholders & help them by adding value to the business workflow with data & analytics
  • )Compute: PySpark, Flink, EMR, EK
  • )Query Engines: Trino, Clickhous
  • eOrchestration: Airflow, Tempora
  • lDevOps: Gitlab, Github Actions, Terrafor
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer II
Data Engineer II

Clear Demand Inc • Chennai District

On-site
INR 1,200,000 - 1,500,000
Data Engineer II
Data Engineer II

CLEAR DEMAND Inc. • Chennai District

On-site
INR 1,200,000 - 1,800,000
Data Engineer I
Data Engineer I

Clear Demand Inc • Chennai District

On-site
INR 1,000,000 - 1,500,000
Data Engineer I
Data Engineer I

CLEAR DEMAND Inc. • Chennai District

On-site
INR 1,200,000 - 1,600,000
Data Engineer
Data Engineer

Advance Career Solutions • Pune District, Chennai District, Bengaluru

Hybrid
INR 1,200,000 - 2,800,000
Data Engineer
Data Engineer

Vandey Global Services • Bengaluru

On-site
INR 800,000 - 1,200,000
Data Engineer – Senior Level
Data Engineer – Senior Level

WebSenor Ltd • Bengaluru, Udaipur District

On-site
INR 1,200,000 - 1,800,000
Data Engineer
Data Engineer

Bounteous • Chennai District

On-site
INR 2,000,000 - 3,600,000
Sr. Data Engineer
Sr. Data Engineer

R Systems • Dadri

On-site
INR 1,200,000 - 1,800,000
Senior Data Engineer
Senior Data Engineer

DATAECONOMY Inc • Hyderabad

On-site
INR 1,500,000 - 2,100,000