Advanced Data Scientist

Antuit India Private Limited

Bengaluru

Hybrid

INR 4,000,000 - 7,000,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work
Flexible hours
Annual well-being day

Job summary

Zebra is hiring a Data Scientist Sr in Bengaluru to lead end-to-end data science projects, build scalable ETL pipelines, and optimize ML/AI models for Retail/CPG use cases. This role interfaces with business stakeholders and operates across cloud platforms like Azure/GCP, Databricks, and PySpark.

The ideal candidate has 6+ years of experience, strong Python/PySpark skills, and hands-on experience with production-grade data and ML pipelines, model explainability, and GenAI/LLMs is a plus, with

Qualifications

  • Master’s degree or relevant work experience in a technical field.
  • 6+ years in Data Science/Data Engineering with end-to-end ML/AI projects.
  • Strong Python/PySpark, SQL and cloud resource management experience.
  • Experience with production-grade data & ML pipelines and model explainability.

Responsibilities

  • Design, optimize, and maintain scalable ETL pipelines on cloud platforms.
  • Develop automated data validation and data quality checks.
  • Build ML/AI models and optimization algorithms; run experiments to show value.
  • Explain data deficiencies and model outputs to business stakeholders.
  • Collaborate with cross-functional teams including Product Management & Software Eng.

Skills

PySpark
Python
SQL
Cloud platforms
Data Pipelines
ML/AI modeling
Model explainability
Communication with business
GenAI/LLMs knowledge

Education

Master’s degree in engineering, computer science, data science, OR relevant field

Tools

Databricks
Git
Airflow
DeltaLiveTables / Unity Catalog
MLFlow

Job description

Overview

At Zebra, we are a community of innovators who come together to create new ways of working. United by curiosity and a culture of caring, we develop smart solutions that anticipate our customer’s and partner’s needs and solve their challenges. Being part of Zebra Nation means you are seen, heard, valued, and respected. Drawing from our unique perspectives, we collaborate to deliver on our purpose. Here you are part of a team pushing boundaries today to redefine the work of tomorrow for organizations, their employees, and those they serve. You’ll have opportunities to learn and lead in a forward-thinking environment, defining your path to a fulfilling career while channeling your skills toward causes you care about—locally and globally. Come make an impact every day at Zebra.

What We're Looking For

The Data Scientist Sr will have experience working across the full lifecycle of a Data Science project. The ideal candidate will demonstrate proficiency in the design, building, and optimization of data ingestion and data transformation pipelines; the design, enhancement and tuning of ML/AI models, the operationalization of resultant models; and communicating root cause analysis and model explainability to business stakeholders. The role will entail interfacing with various source systems, and proficiency in PySpark, SQL, Databricks, and related cloud resource management services. Prior experience with advanced analytics, ML/AI, and optimization solutions in the Retail/CPG domain is strongly preferred. A working knowledge of GenAI/LLMs and Agentic-AI in the context of workflow automation will be valuable but is not the primary requirement for this position.

Responsibilities
  • Design, optimize, and maintain scalable ETL pipelines using PySpark and Databricks on cloud platforms (Azure/GCP).
  • Develop automated data validation process to proactively perform data quality checks.
  • Employ key Databricks modules (DeltaLiveTables, Unity Catalog, MLFlow) to facilitate creating, running experiments, automating, and scheduling jobs on Databricks.
  • Optimize the allocation of cloud resources and Databricks DBUs to manage and control cloud and Databricks consumption cost.
  • Employ GitHub repositories to ensure that production code management best practices are being strictly adhered to.
  • Build and tune ML/AI and optimization models, identify algorithmic performance improvement opportunities, and perform experiments to demonstrate incremental value delivered.
  • Have frequent conversations with Business Stakeholders to understand their requirements and concerns.
  • Explain data deficiencies, model performance/root cause analysis, and model output.
  • Follow best practices in Data Architecture, Coding, and Project Management operations.
  • Collaborate with cross-functional teams, such as Customers’ Stakeholders, Engagement Managers, Data Ops/Job Monitoring, Product Management & Software Engineering.
  • Expand the use of analytics, ML/AI, mathematical optimization, Gen-AI/LLMs and Agentic-AI in the context of Retail/CPG business use cases such as anomaly detection, demand forecasting, price elasticity modeling, promotions features & strategy simulation, product cannibalization and halo modeling, markdown optimization, product allocation, reorder/replenishment, size and pack optimization, workforce scheduling & task optimization.
Qualifications
Job Requirements

Minimum Education: Master’s degree in engineering, computer science, data science, operations research, statistics, mathematics, quantitative sciences or relevant work experience.

Minimum Work Experience (years): 6+ years of experience in Data Science/Data Engineering with emphasis on the full lifecycle of Data Science-ML/AI projects. Within that timeframe, experience is expected in: Python/PySpark, SQL, and relational or NoSQL databases, and cloud resource management.

Key Skills and Competencies
  • Experience working with AWS, Azure, or GCP cloud environments.
  • Experience implementing advanced analytics, ML/AI algorithms (such as: , Statistical Time Series: Exponential Smoothing Models, S/ARIMA; Machine Learning: Random Forests, Gradient Boosting Methods; Neural Networks: TiDE (Google), DenseNet & Prophet (FB-Meta); Foundational Time Series Models: TimesFM (Google), Chronos (AWS) Mathematical (constrained linear, non-linear and network) optimization models
  • Proven experience building end-to-end production grade Data & ML/AI pipelines using PySpark, Python (Pandas/NumPy) and SQL.
  • Experience working with Git (or similar code management repositories) as a collaboration tool. production-grade
  • Experience with orchestration tools like Databricks, Airflow (or similar tools like Snowflake, Dagster, etc.).
  • Understanding of Retail/CPG industry business challenges with an emphasis on Supply Chain, Pricing/Promotions/Markdowns, Inventory Allocation, Assortment Mix Planning, Size & Pack Optimization, Retail Shrink/Fraud & Anomaly Detection, and Workforce optimization applications are highly desirable.
  • Excellent verbal and written communication skills, especially as it relates to technical communications.
  • Ability to present technical analysis to business stakeholders.
  • Demonstrated ability to learn new technologies quickly and independently, particularly as technology in this domain rapidly advances.
  • In this context, a working knowledge of GenAI/LLMs, Agentic-AI and related frameworks (e.g. LangChain) will be a plus.
  • Ability to work independently with minimal supervision and achieve stretch goals in a highly innovative and a fast-paced environment.
Benefits

We understand the importance of work-life balance and wellbeing, which is why we offer flexibility for our teams including: hybrid work, adaptable hours, Summer Flex Fridays, Focus Fridays, and an annual companywide well-being day to promote revitalization and success.

Zebra provides the foundation for intelligent operations with an award-winning portfolio of connected frontline, asset visibility and automation solutions.

Organizations globally across retail, manufacturing, transportation, logistics, healthcare, and other industries rely on us to deliver outcomes today while driving innovation for what's next.

Together with our partners, we create new ways of working that improve productivity and empower organizations to be better every day.

Learn more at zebra.com.

Zebra Better Every Day

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Advanced Data Scientist
Advanced Data Scientist

Zebra Technologies • Bengaluru

Hybrid
INR 1,200,000 - 1,800,000
Hybrid work
Flexible hours
Summer Flex Fridays
+2
Supervisor, Software Engineering (IN)
Supervisor, Software Engineering (IN)

Antuit India Private Limited • Pune District

On-site
INR 3,000,000 - 6,000,000
Hybrid work option
Adaptable hours
Summer Flex Fridays
+2
Data Scientist/ Sr. Data Scientist
Data Scientist/ Sr. Data Scientist

Jaspercolin • India

On-site
INR 700,000 - 1,500,000
Attractive salary
Benefits and performance-based incentives
Continuous learning opportunities
+1
Senior Analyst-Data Science
Senior Analyst-Data Science

AMERICAN EXPRESS • Sultanpur

On-site
INR 900,000 - 1,300,000
Sr. Data Scientist / Data Scientist
Sr. Data Scientist / Data Scientist

Williams-Sonoma, Inc. • Pune District

On-site
INR 1,500,000 - 2,300,000
Lead, Data Scientist [T500-17878]
Lead, Data Scientist [T500-17878]

ANSR • Bengaluru

On-site
INR 1,500,000 - 2,500,000
Data Scientist-Advanced Analytics
Data Scientist-Advanced Analytics

IBM • Dadri, Gurugram District

On-site
INR 1,500,000 - 3,000,000
Data Scientist - Senior
Data Scientist - Senior

PowerToFly • Pune District

On-site
INR 1,500,000 - 2,500,000
Sr. / Data Scientist
Sr. / Data Scientist

Williams-Sonoma, Inc. • Pune District

On-site
INR 1,500,000 - 3,000,000
Software Engineer Senior I
Software Engineer Senior I

Zebra Technologies • Bengaluru

Hybrid
INR 1,800,000 - 2,800,000
Hybrid work
Summer Flex Fridays
Focus Fridays
+1