Data Engineer II

Visa Hunt

India

On-site

INR 1,200,000 - 1,800,000

Full time

6 days ago
Be an early applicant
Application generator

Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.

Get past ATS filters

Benefits offered by this job

Remote environments
Learning stipend
Autonomy and mentorship
Competitive compensation

Job summary

Mrsool is seeking a Data Engineer II to build and scale a data platform that powers analytics, product insights, and data-driven decisions across the organization. You’ll work with Kafka, Maxwell, Spark, S3, Trino, BigQuery, dbt and Metabase to construct reliable, scalable data systems and trusted datasets for Product, Analytics, Data Science and Business teams.

You’ll design and implement robust data pipelines, modern data lake architectures, reusable data models, and prioritize performance,

Qualifications

  • 4+ years of hands-on experience designing and building scalable data platforms, data lakes, and data warehouses.
  • Strong proficiency in Spark (Scala, python) and SQL, with experience building production-grade data pipelines and distributed data processing applications.
  • Hands‑on experience with Apache Spark and a solid understanding of distributed data processing, performance tuning, and optimization.
  • Experience building batch and streaming data pipelines using technologies such as Kafka, CDC/Maxwell, or similar event-driven architectures.
  • Strong understanding of modern data lake architectures, including Medallion Architecture, data modeling, partitioning, and storage optimization.
  • Experience working with cloud-native data platforms and technologies such as Amazon S3, BigQuery, Trino, or similar analytics engines.
  • Solid experience designing dimensional models, star schemas, and building reliable data marts that support analytics and business intelligence.
  • Hands‑on experience with dbt, including developing reusable models, implementing automated testing, and maintaining documentation.
  • Strong knowledge of data quality, observability, lineage, and engineering best practices to build reliable and maintainable data products.
  • Experience optimizing large-scale data pipelines, SQL queries, and distributed processing jobs for performance, scalability, and cost efficiency.
  • Familiarity with CI/CD, Git-based development workflows, infrastructure automation, and modern software engineering best practices.
  • Excellent problem‑solving skills with the ability to independently own projects from design through production.
  • Strong communication and stakeholder management skills, with experience collaborating across Product, Engineering, Analytics, and Business teams.
  • A passion for building scalable data platforms and continuously improving developer experience, platform reliability, and operational excellence.

Responsibilities

  • Design, build, and maintain scalable batch and real-time data pipelines using Maxwell, Kafka, Spark, and dbt to power analytics and business-critical applications.
  • Develop and optimize data models following Medallion Architecture (Bronze, Silver, Gold) to create reliable, reusable, and high-quality datasets.
  • Build and maintain cloud-native data platforms using S3, Spark, Trino, and BigQuery, ensuring scalability, reliability, and cost efficiency.
  • Design robust data ingestion frameworks leveraging CDC (Maxwell), Kafka, and event-driven architectures to support near real-time data processing.
  • Create, optimize, and maintain data warehouses and data marts that enable fast, reliable reporting and self-service analytics.
  • Partner closely with Product Managers, Data Analysts, Backend Engineers, and Business stakeholders to translate business requirements into scalable data solutions.
  • Develop reusable dbt models, testing frameworks, and documentation to improve data quality, governance, and developer productivity.
  • Optimize Spark jobs, Trino queries, and storage layouts for performance, reliability, and cost efficiency.
  • Own the end-to-end lifecycle of critical data pipelines, ensuring high availability, monitoring, SLA adherence, and proactive incident resolution.
  • Build and enhance the core data platform by developing reusable frameworks, automation, CI/CD pipelines, and engineering best practices.
  • Ensure data quality through validation, monitoring, lineage, and observability while implementing best practices for security and governance.
  • Enable analytics teams by delivering trusted datasets, semantic models, and dashboards that power decision-making through Metabase.

Skills

Spark
SQL
Data pipelines
Data modeling
Cloud platforms
Stakeholder management

Tools

Kafka
Maxwell
dbt
S3
BigQuery
Trino
Metabase

Job description

Welcome to the world of Mrsool! Where on-demand delivery meets unparalleled user needs to deliver anything you desire. As one of the largest delivery platforms in the Middle East and North Africa (MENA) region, Mrsool has captivated users with its unique and seamless experience, earning it the highest ratings among all major delivery platforms on both Apple's App Store and Google's Play Store.

What sets Mrsool apart is its commitment to providing an unmatched "order anything from anywhere" experience. This extraordinary feat is made possible by our extensive fleet of dedicated on-demand couriers. With their unwavering dedication, they ensure that your desired items reach your doorstep, no matter where you are.

Whether it's a late-night craving, a forgotten item, or a special gift for a loved one, Mrsool is here to deliver, quite literally. We take pride in the convenience we offer, empowering you to get what you need when you need it, all at the tap of a button.

The Job in a Nutshell

We’re looking for a passionate Data Engineer II to help build and scale the data platform that powers analytics, product insights, and data-driven decision making across the organization. You’ll work on modern data technologies including Kafka, Maxwell, Spark, S3, Trino, BigQuery, dbt, and Metabase to build reliable, scalable, and high-performance data systems. From real-time data ingestion and distributed data processing to dimensional modeling and data platform development, you’ll play a key role in shaping the foundation of our data ecosystem.

As part of the Data Engineering team, you’ll design and build robust data pipelines, implement modern data lake architectures, develop reusable data models, and create trusted datasets that empower Product, Analytics, Data Science, and Business teams. You’ll also contribute to improving our engineering standards by focusing on performance, reliability, observability, automation, and developer experience.

This role is ideal for someone who enjoys solving complex data engineering challenges, building scalable platforms, and continuously improving how data is collected, transformed, and consumed across the organization.

If you’re excited about building modern data platforms, working with large-scale distributed systems, and having a meaningful impact on the company’s data strategy, we’d love to hear from you.

What You Will Do
  • Design, build, and maintain scalable batch and real-time data pipelines using Maxwell, Kafka, Spark, and dbt to power analytics and business-critical applications.
  • Develop and optimize data models following Medallion Architecture (Bronze, Silver, Gold) to create reliable, reusable, and high-quality datasets.
  • Build and maintain cloud-native data platforms using S3, Spark, Trino, and BigQuery, ensuring scalability, reliability, and cost efficiency.
  • Design robust data ingestion frameworks leveraging CDC (Maxwell), Kafka, and event-driven architectures to support near real-time data processing.
  • Create, optimize, and maintain data warehouses and data marts that enable fast, reliable reporting and self-service analytics.
  • Partner closely with Product Managers, Data Analysts, Backend Engineers, and Business stakeholders to translate business requirements into scalable data solutions.
  • Develop reusable dbt models, testing frameworks, and documentation to improve data quality, governance, and developer productivity.
  • Optimize Spark jobs, Trino queries, and storage layouts for performance, reliability, and cost efficiency.
  • Own the end-to-end lifecycle of critical data pipelines, ensuring high availability, monitoring, SLA adherence, and proactive incident resolution.
  • Build and enhance the core data platform by developing reusable frameworks, automation, CI/CD pipelines, and engineering best practices.
  • Ensure data quality through validation, monitoring, lineage, and observability while implementing best practices for security and governance.
  • Enable analytics teams by delivering trusted datasets, semantic models, and dashboards that power decision-making through Metabase.
Requirements
  • 4+ years of hands-on experience designing and building scalable data platforms, data lakes, and data warehouses.
  • Strong proficiency in Spark (Scala, python) and SQL, with experience building production-grade data pipelines and distributed data processing applications.
  • Hands‑on experience with Apache Spark and a solid understanding of distributed data processing, performance tuning, and optimization.
  • Experience building batch and streaming data pipelines using technologies such as Kafka, CDC/Maxwell, or similar event-driven architectures.
  • Strong understanding of modern data lake architectures, including Medallion Architecture, data modeling, partitioning, and storage optimization.
  • Experience working with cloud-native data platforms and technologies such as Amazon S3, BigQuery, Trino, or similar analytics engines.
  • Solid experience designing dimensional models, star schemas, and building reliable data marts that support analytics and business intelligence.
  • Hands‑on experience with dbt, including developing reusable models, implementing automated testing, and maintaining documentation.
  • Strong knowledge of data quality, observability, lineage, and engineering best practices to build reliable and maintainable data products.
  • Experience optimizing large-scale data pipelines, SQL queries, and distributed processing jobs for performance, scalability, and cost efficiency.
  • Familiarity with CI/CD, Git-based development workflows, infrastructure automation, and modern software engineering best practices.
  • Excellent problem‑solving skills with the ability to independently own projects from design through production.
  • Strong communication and stakeholder management skills, with experience collaborating across Product, Engineering, Analytics, and Business teams.
  • A passion for building scalable data platforms and continuously improving developer experience, platform reliability, and operational excellence.
Benefits
What We Offer You

Inclusive and Diverse Environment: We foster an inclusive and diverse workplace that values innovation and offers remote environments.

Competitive Compensation: Our compensation packages are highly competitive and include potential share options for certain roles.

Personal Growth and Development: We are committed to your personal and professional growth, providing regular training and an annual learning stipend to help you advance your career in a dynamic environment.

Autonomy and Mentorship: You'll enjoy a high degree of autonomy in your role, supported by mentorship and ambitious goals that pave the way for both your success and the company's growth.

Originally posted on Himalayas

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Scientist II
Data Scientist II

Mrsool • India

On-site
INR 1,800,000 - 3,200,000
Remote-friendly environment
Share options
Annual learning stipend
+1
Sr. Data Engineer
Sr. Data Engineer

Kanerika Software Pvt Ltd. • Indore District

On-site
INR 1,200,000 - 1,800,000
Health Insurance
Flexible Working Hours
Professional Certification Reimbursements
+2
Sr. Data Engineer
Sr. Data Engineer

Cohere Health • Hyderabad

On-site
INR 1,200,000 - 1,800,000
Data Engineering Engineer II
Data Engineering Engineer II

TekWissen India • Chennai District

Hybrid
INR 1,800,000 - 2,400,000
Software Engineer II
Software Engineer II

Metlife • Hyderabad

Hybrid
INR 1,500,000 - 3,000,000
Data Engineer ( 2 to 5 Years)
Data Engineer ( 2 to 5 Years)

Siemens Mobility • Bengaluru

Hybrid
INR 1,500,000 - 2,600,000
Hybrid working
Diverse and collaborative culture
Learning and development opportunities
+1
Senior Data Engineer (Data & Analytics)
Senior Data Engineer (Data & Analytics)

Summit Consulting Services • Ernakulam

On-site
INR 1,000,000 - 1,800,000
Collaborative culture
Opportunity to influence architecture
Modern cloud-native tech stack
Lead Engineer - Data Engineering
Lead Engineer - Data Engineering

Pine Labs • Dadri

On-site
INR 4,000,000 - 7,500,000
Senior Data Engineer
Senior Data Engineer

o9 Solutions, Inc. • Bengaluru

Hybrid
INR 900,000 - 1,300,000
Senior Software Engineer II - Data Engineering & Platform
Senior Software Engineer II - Data Engineering & Platform

Playsimple Games Private Limited • Bengaluru

Hybrid
INR 4,000,000 - 7,000,000