Data Scientist: Hyderabad and Gurugram
You will be part of a collaborative interdisciplinary team around data, where you will be responsible for our continuous delivery of statistical/ML models. You will work closely with process owners, product owners, and final business users. This will provide you the correct visibility and understanding of the criticality of your developments.
Responsibilities
- Delivery of key Advanced Analytics/Data Science projects within time and budget, particularly around DevOps/MLOps and Machine Learning models in scope.
- Active contributor to code & development in projects and services.
- Partner with data engineers to ensure data access for discovery and proper data is prepared for model consumption.
- Partner with ML engineers working on industrialization.
- Communicate with business stakeholders in the process of service design, training, and knowledge transfer.
- Support large-scale experimentation and build data-driven models.
- Refine requirements into modelling problems.
- Influence product teams through data-based recommendations.
- Research in state-of-the-art methodologies.
- Create documentation for learnings and knowledge transfer.
- Create reusable packages or libraries.
- Ensure on-time and on-budget delivery which satisfies project requirements while adhering to enterprise architecture standards.
- Leverage big data technologies to help process data and build scaled data pipelines (batch to real-time).
- Implement end-to-end ML lifecycle with Azure Machine Learning and Azure Pipelines.
Qualifications
- BE/B.Tech in Computer Science, Maths, or technical fields.
- Overall 5+ years of experience working as a Data Scientist.
- 4+ years’ experience building solutions in the commercial or supply chain space.
- 4+ years working in a team to deliver production-level analytic solutions. Fluent in git (version control). Understanding of Jenkins and Docker are a plus.
- Fluent in SQL syntax.
- 4+ years’ experience in Statistical/ML techniques to solve supervised (regression, classification) and unsupervised problems.
- 4+ years’ experience in developing business problem-related statistical/ML modeling with industry tools with primary focus on Python or Pyspark development.
Skills, Abilities, Knowledge:
- Data Science – Hands-on experience and strong knowledge of building machine learning models – supervised and unsupervised models. Knowledge of Time series/Demand Forecast models is a plus.
- Programming Skills – Hands-on experience in statistical programming languages like Python, Pyspark and database query languages like SQL.
- Statistics – Good applied statistical skills, including knowledge of statistical tests, distributions, regression, maximum likelihood estimators.
- Cloud (Azure) – Experience in Databricks and ADF is desirable.
- Familiarity with Spark, Hive, Pig is an added advantage.
- Business storytelling and communicating data insights in a business consumable format. Fluent in one Visualization tool.
- Strong communications and organizational skills with the ability to deal with ambiguity while juggling multiple priorities.
- Experience with Agile methodology for teamwork and analytics ‘product’ creation.
Seniority level
Mid-Senior level
Employment type
Full-time
Job function
Information Technology
Industries
Food and Beverage Services, IT Services and IT Consulting, and Financial Services