Data Engineer II

Mrsool 3

India

Remote

INR 1.800.000 - 3.000.000

Vollzeit

14 Tage+
Bewerbungsgenerator

Bekomme eine Antwort von diesem Arbeitgeber — ein Lebenslauf und ein Anschreiben, die genau auf die Eigenschaften eingehen, die gesucht werden.

Schaffe es an den ATS-Filtern vorbei

Benefits dieser Stelle

Remote-friendly environment
Annual learning stipend
Competitive compensation

Zusammenfassung

Mrsool is seeking a Data Engineer II to design and scale a modern data platform that powers analytics, product insights, and data-driven decision making across the organization. You will work with Kafka, Maxwell, Spark, S3, Trino, BigQuery, dbt, and Metabase to build reliable data pipelines and reusable models, enabling cross-team data collaboration.

This role emphasizes performance, observability, and scalable data architectures, with opportunities to advance data strategy and engineering

Qualifikationen

  • 4+ years of hands-on experience designing and building scalable data platforms, data lakes, and data warehouses.
  • Strong proficiency in Spark (Scala, Python) and SQL, with experience building production-grade data pipelines and distributed data processing applications.
  • Hands-on experience with Apache Spark and a solid understanding of distributed data processing, performance tuning, and optimization.
  • Experience building batch and streaming data pipelines using technologies such as Kafka, CDC/Maxwell, or similar event-driven architectures.
  • Strong understanding of modern data lake architectures, including Medallion Architecture, data modeling, partitioning, and storage optimization.
  • Experience working with cloud-native data platforms and technologies such as Amazon S3, BigQuery, Trino, or similar analytics engines.
  • Solid experience designing dimensional models, star schemas, and building reliable data marts that support analytics and business intelligence.
  • Hands-on experience with dbt, including developing reusable models, implementing automated testing, and maintaining documentation.
  • Strong knowledge of data quality, observability, lineage, and engineering best practices to build reliable and maintainable data products.
  • Experience optimizing large-scale data pipelines, SQL queries, and distributed processing jobs for performance, scalability, and cost efficiency.
  • Familiarity with CI/CD, Git-based development workflows, infrastructure automation, and modern software engineering best practices.
  • Excellent problem-solving skills with the ability to independently own projects from design through production.
  • Strong communication and stakeholder management skills, with experience collaborating across Product, Engineering, Analytics, and Business teams.
  • A passion for building scalable data platforms and continuously improving developer experience, platform reliability, and operational excellence.

Aufgaben

  • Design, build, and maintain scalable batch and real-time data pipelines using Maxwell, Kafka, Spark, and dbt to power analytics and business-critical applications.
  • Develop and optimize data models following Medallion Architecture (Bronze, Silver, Gold) to create reliable, reusable, and high-quality datasets.
  • Build and maintain cloud-native data platforms using S3, Spark, Trino, and BigQuery, ensuring scalability, reliability, and cost efficiency.
  • Design robust data ingestion frameworks leveraging CDC (Maxwell), Kafka, and event-driven architectures to support near real-time data processing.
  • Create, optimize, and maintain data warehouses and data marts that enable fast, reliable reporting and self-service analytics.
  • Partner closely with Product Managers, Data Analysts, Backend Engineers, and Business stakeholders to translate business requirements into scalable data solutions.
  • Develop reusable dbt models, testing frameworks, and documentation to improve data quality, governance, and developer productivity.
  • Optimize Spark jobs, Trino queries, and storage layouts for performance, reliability, and cost efficiency.
  • Own the end-to-end lifecycle of critical data pipelines, ensuring high availability, monitoring, SLA adherence, and proactive incident resolution.
  • Build and enhance the core data platform by developing reusable frameworks, automation, CI/CD pipelines, and engineering best practices.
  • Ensure data quality through validation, monitoring, lineage, and observability while implementing best practices for security and governance.
  • Enable analytics teams by delivering trusted datasets, semantic models, and dashboards that power decision-making through Metabase.

Kenntnisse

Spark
SQL
Kafka
dbt
Medallion Arch
BigQuery
Trino
S3
Metabase
Python

Tools

Kafka
Maxwell
Spark
dbt
S3
Trino
BigQuery
Metabase

Jobbeschreibung

Who Are We

Welcome to the world of Mrsool! Where on-demand delivery meets unparalleled user needs to deliver anything you desire. As one of the largest delivery platforms in the Middle East and North Africa (MENA) region, Mrsool has captivated users with its unique and seamless experience, earning it the highest ratings among all major delivery platforms on both Apple's App Store and Google's Play Store.


What sets Mrsool apart is its commitment to providing an unmatched \"order anything from anywhere\" experience. This extraordinary feat is made possible by our extensive fleet of dedicated on-demand couriers. With their unwavering dedication, they ensure that your desired items reach your doorstep, no matter where you are.


Whether it's a late-night craving, a forgotten item, or a special gift for a loved one, Mrsool is here to deliver, quite literally. We take pride in the convenience we offer, empowering you to get what you need when you need it, all at the tap of a button.


The Job in a Nutshell


We're looking for a passionate Data Engineer II to help build and scale the data platform that powers analytics, product insights, and data-driven decision making across the organization.


You’ll work on modern data technologies including Kafka, Maxwell, Spark, S3, Trino, BigQuery, dbt, and Metabase to build reliable, scalable, and high-performance data systems. From real-time data ingestion and distributed data processing to dimensional modeling and data platform development, you’ll play a key role in shaping the foundation of our data ecosystem.


As part of the Data Engineering team, you’ll design and build robust data pipelines, implement modern data lake architectures, develop reusable data models, and create trusted datasets that empower Product, Analytics, Data Science, and Business teams. You’ll also contribute to improving our engineering standards by focusing on performance, reliability, observability, automation, and developer experience.


This role is ideal for someone who enjoys solving complex data engineering challenges, building scalable platforms, and continuously improving how data is collected, transformed, and consumed across the organization.


If you’re excited about building modern data platforms, working with large-scale distributed systems, and having a meaningful impact on the company’s data strategy, we’d love to hear from you.


What You Will Do


  • Design, build, and maintain scalable batch and real-time data pipelines using Maxwell, Kafka, Spark, and dbt to power analytics and business-critical applications.

  • Develop and optimize data models following Medallion Architecture (Bronze, Silver, Gold) to create reliable, reusable, and high-quality datasets.

  • Build and maintain cloud-native data platforms using S3, Spark, Trino, and BigQuery, ensuring scalability, reliability, and cost efficiency.

  • Design robust data ingestion frameworks leveraging CDC (Maxwell), Kafka, and event-driven architectures to support near real-time data processing.

  • Create, optimize, and maintain data warehouses and data marts that enable fast, reliable reporting and self-service analytics.

  • Partner closely with Product Managers, Data Analysts, Backend Engineers, and Business stakeholders to translate business requirements into scalable data solutions.

  • Develop reusable dbt models, testing frameworks, and documentation to improve data quality, governance, and developer productivity.

  • Optimize Spark jobs, Trino queries, and storage layouts for performance, reliability, and cost efficiency.

  • Own the end-to-end lifecycle of critical data pipelines, ensuring high availability, monitoring, SLA adherence, and proactive incident resolution.

  • Build and enhance the core data platform by developing reusable frameworks, automation, CI/CD pipelines, and engineering best practices.

  • Ensure data quality through validation, monitoring, lineage, and observability while implementing best practices for security and governance.

  • Enable analytics teams by delivering trusted datasets, semantic models, and dashboards that power decision-making through Metabase.


Requirements

What We’re Looking For


  • 4+ years of hands-on experience designing and building scalable data platforms, data lakes, and data warehouses.

  • Strong proficiency in Spark (Scala, python) and SQL, with experience building production-grade data pipelines and distributed data processing applications.

  • Hands-on experience with Apache Spark and a solid understanding of distributed data processing, performance tuning, and optimization.

  • Experience building batch and streaming data pipelines using technologies such as Kafka, CDC/Maxwell, or similar event-driven architectures.

  • Strong understanding of modern data lake architectures, including Medallion Architecture, data modeling, partitioning, and storage optimization.

  • Experience working with cloud-native data platforms and technologies such as Amazon S3, BigQuery, Trino, or similar analytics engines.

  • Solid experience designing dimensional models, star schemas, and building reliable data marts that support analytics and business intelligence.

  • Hands-on experience with dbt, including developing reusable models, implementing automated testing, and maintaining documentation.

  • Strong knowledge of data quality, observability, lineage, and engineering best practices to build reliable and maintainable data products.

  • Experience optimizing large-scale data pipelines, SQL queries, and distributed processing jobs for performance, scalability, and cost efficiency.

  • Familiarity with CI/CD, Git-based development workflows, infrastructure automation, and modern software engineering best practices.

  • Excellent problem-solving skills with the ability to independently own projects from design through production.

  • Strong communication and stakeholder management skills, with experience collaborating across Product, Engineering, Analytics, and Business teams.

  • A passion for building scalable data platforms and continuously improving developer experience, platform reliability, and operational excellence.


Benefits

What We Offer You

Inclusive and Diverse Environment: We foster an inclusive and diverse workplace that values innovation and offers remote environments.


Competitive Compensation: Our compensation packages are highly competitive and include potential share options for certain roles.


Personal Growth and Development: We are committed to your personal and professional growth, providing regular training and an annual learning stipend to help you advance your career in a dynamic environment.


Autonomy and Mentorship: You'll enjoy a high degree of autonomy in your role, supported by mentorship and ambitious goals that pave the way for both your success and the company’s growth.

Hol dir deinen kostenlosen, vertraulichen Lebenslauf-Check.

oder ziehe deine Datei hierhin.

Similar jobs

Ähnliche Jobs, die dir auch gefallen könnten

Engineering Manager
Engineering Manager

Mrsool 3 • Indien

Remote
INR 700.000 - 1.200.000
Inclusive and Diverse Environment
Competitive compensation with share/RS
Learning stipend and growth support
+1
Data Scientist II
Data Scientist II

Mrsool 3 • Indien

Remote
INR 1.500.000 - 2.800.000
Remote-friendly environment
Annual learning stipend
Regular training
Data Engineer - Senior II
Data Engineer - Senior II

Zoho • New Delhi

Vor Ort
INR 3.500.000 - 7.000.000
Flexible work arrangements
Comprehensive insurance coverage
SDE III Backend Engineer
SDE III Backend Engineer

Mrsool 3 • Indien

Hybrid
INR 2.500.000 - 4.000.000
Remote-friendly environment
Share options
Learning stipend
+1
Analyst , Performance
Analyst , Performance

Mrsool 3 • Indien

Vor Ort
INR 1.200.000 - 1.800.000
Inclusive environment
Competitive compensation
Annual learning stipend
+2
Site Reliability Engineer II
Site Reliability Engineer II

Mrsool 3 • Indien

Remote
INR 1.800.000 - 2.400.000
Remote environments
Competitive compensation
Annual learning stipend
+1
Senior Data Engineer
Senior Data Engineer

UNAVAILABLE • Bengaluru

Vor Ort
INR 4.000.000 - 7.000.000
Annual Leave 15 days
Sick Leave 12 days
Parental Leave
+3
Data Engineer
Data Engineer

Bayzat • Delhi

Vor Ort
INR 1.500.000 - 2.100.000
Remote and Hybrid options
Sr. Data Engineer
Sr. Data Engineer

Kanerika Software Pvt Ltd. • Indore District

Vor Ort
INR 1.200.000 - 1.800.000
Health Insurance
Flexible Working Hours
Professional Certification Reimbursements
+2
Senior Data Engineer (Data & Analytics)
Senior Data Engineer (Data & Analytics)

Summit Consulting Services • Ernakulam

Vor Ort
INR 1.000.000 - 1.800.000
Collaborative culture
Opportunity to influence architecture
Modern cloud-native tech stack