Data Scientist II

Keka Inc.

Karnataka

On-site

INR 1,500,000 - 3,000,000

Full time

31 hours ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Skan AI is seeking an experienced Data Scientist who loves hands-on work to build and own AI/ML models and data pipelines. You will work with large real-world datasets and be the go-to person when something breaks or doesn't make sense.

You will explore DL, NLP, time-series, and process mining; write production-grade Python; design robust data pipelines; ensure data quality; and collaborate with engineering to optimize ingestion and feature generation for speed and reliability.

Qualifications

  • MS or B.Tech/B.E. in CS/Stats/Math or related quantitative field, or equivalent hands-on experience.
  • 3–7 years of industry experience building and shipping DS/ML in production.
  • Strong Python skills with clean, modular, testable code.
  • Experience with large-scale data processing using Spark/Databricks or similar frameworks.
  • Proficiency in at least two of: supervised/unsupervised ML, NLP, DL, sequence modelling, or process mining.
  • Experience diagnosing and fixing real-world data quality and model issues in production.

Responsibilities

  • Build, train, evaluate, and improve ML models powering core features.
  • Own the full modelling lifecycle: data exploration, feature engineering, model selection, tuning, validation.
  • Run disciplined experiments with clear hypotheses and evaluation frameworks.
  • Explore DL, NLP, time-series, and unsupervised learning for process intelligence problems.
  • Write production-grade Python code and build reliable data pipelines for multimodal enterprise data.
  • Ensure data quality, monitor pipelines, and fix issues proactively.

Skills

Python programming
Spark/Databricks
ML modelling
Data pipelines
Communication
Problem solving

Education

MS or B.Tech/B.E.

Tools

Spark
Databricks
Python

Job description

Be at the Forefront of the Agentic AI Revolution

At Skan AI, we are pioneering the context engine for human and agentic execution, bringing context from enterprise operators, systems, and processes to power how the world's largest organizations execute their most complex, mission-critical work.

We're in hyper-growth mode at exactly the right moment in history. As enterprises race to adopt agentic AI, we're uniquely positioned to deliver the clear signal they desperately need: a platform that trains and grounds AI Agents in trillions of real execution signals, enabling reliable, compliant automation of their most complex processes.

Backed by Dell Technologies Capital and other leading investors, we're the only company that can bridge the gap between AI's promise and enterprise reality, making us perfectly positioned to define the agentic era for modern enterprises.

Our diverse, collaborative team of 250+ innovators is solving category-defining challenges at the intersection of AI, process intelligence, and enterprise work. Diverse perspectives fuel breakthrough thinking, cross-functional collaboration is the norm, and our work directly transforms how Fortune 500 companies operate. We are shaping the future of work itself.

The Role

We are looking for an experienced Data Scientist who loves to get their hands dirty. This is a deeply hands-on role; you will build and own AI/ML models and data pipelines, work directly with large and messy real-world datasets, and be the go-to person when something breaks or doesn't make sense.

What You'll Do
Hands-On Modelling & Algorithm Development
  • Build, train, evaluate, and improve ML models that power core Skan.ai features including but not limited to task detection, workflow segmentation, behavioral pattern recognition, process discovery, and variant analysis.
  • Own the full modelling lifecycle: data exploration, feature engineering, model selection, hyperparameter tuning and validation.
  • Run disciplined experiments: define clear hypotheses, design evaluation frameworks, and present findings with statistical rigour.
  • Explore and apply techniques from deep learning, NLP, time-series analysis, and unsupervised learning to real process intelligence problems.
  • Write clean, well-tested, production-ready Python code that your teammates can maintain and build on.
  • Design and build reliable data pipelines that handle high volumes of multimodal enterprise data such as screen telemetry, event logs, clickstreams, and structured process data.
  • Own data quality at every step: define validation checks, detect drift, monitor pipeline health, and fix issues proactively.
  • Collaborate with engineering to optimise ingestion and feature generation workflows for speed, cost, and reliability.
Customer Issue Diagnosis & Production Support
  • Investigate and resolve data and model issues that surface in customer environments; trace root causes across pipelines, features, and model behaviour.
  • Partner with Customer Success and Solutions teams to understand unexpected outcomes and translate them into concrete technical fixes.
Cross-Functional Collaboration
  • Work closely with product and engineering to understand requirements, scope data work, and deliver on time.
  • Communicate findings clearly to non-technical stakeholders; translate model outputs and data insights into plain language that drives decisions.
  • Contribute technical input to sprint planning and quarterly priorities; flag feasibility concerns early and propose alternatives.
What We’re Looking For
Required
  • M.S. or B.Tech/B.E. in Computer Science, Statistics, Mathematics, or a related quantitative field or equivalent hands-on experience.
  • 3-7 years of industry experience building and shipping data science or ML solutions in production.
  • Strong Python skills with clean, modular, testable code as a baseline expectation.
  • Solid experience with large-scale data processing using Spark, Databricks, or similar distributed frameworks.
  • Proficiency in at least two of: supervised/unsupervised ML, NLP, deep learning, sequence modelling, or process mining.
  • Experience diagnosing and fixing real-world data quality and model behaviour issues in production systems.
  • Clear, concise communication: you can explain a complex model or a data issue to an engineer, a PM, or a customer success manager without jargon.
Nice to Have
  • Exposure to process mining, RPA, workflow analytics, or enterprise operations intelligence.
  • Familiarity with LLMs, transformer-based models, or multimodal learning pipelines.
  • Prior experience in a product-focused startup or scale-up environment.

Skan AI is an equal opportunity employer committed to building a diverse, inclusive, and respectful workplace around the world. We do not discriminate based on race, color, religion or belief, sex (including pregnancy, sexual orientation, gender identity, or gender expression), national origin, ancestry, age, disability, medical condition, genetic information, marital or family status, military or veteran status, or any other characteristic protected by applicable laws in the locations where we operate.

We welcome people from all backgrounds and provide reasonable accommodations throughout the hiring process.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Sr. Data Scientist
Sr. Data Scientist

Keka Inc. • Karnataka

On-site
INR 4,000,000 - 7,000,000
Solution Architect
Solution Architect

Skanai • India

On-site
INR 350,000 - 700,000
Lead AI Solutions Consultant
Lead AI Solutions Consultant

Skanai • India

On-site
INR 3,000,000 - 5,500,000
Lead AI Solutions Consultant
Lead AI Solutions Consultant

SkanAI • Karwar

On-site
INR 4,000,000 - 7,000,000
AI Solutions Consultant
AI Solutions Consultant

Skan, Inc. • India

On-site
INR 1,200,000 - 1,800,000
Principal Product Marketing Manager, Agentic AI
Principal Product Marketing Manager, Agentic AI

SkanAI • Karnataka

On-site
INR 13,384,000 - 20,076,000
Lead DevOps Engineer
Lead DevOps Engineer

SkanAI • Karnataka

Hybrid
INR 2,800,000 - 4,000,000
Hybrid work model
Professional growth
Principal Product Marketing Manager, Agentic AI
Principal Product Marketing Manager, Agentic AI

Skan • India

Remote
INR 3,000,000 - 6,000,000
Enterprise Account Executive, Banking New York
Enterprise Account Executive, Banking New York

Skanai • India

On-site
USD 150,000 - 200,000
Competitive compensation
Uncapped variable plan
Equity participation
+2
Lead Software Development Engineer (Python)
Lead Software Development Engineer (Python)

Keka Inc. • Karnataka

Hybrid
INR 6,000,000 - 9,000,000