Data Scientist - Senior

JobCubby

Pune District

On-site

INR 3,000,000 - 6,000,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

JobCubby is seeking an experienced Data Platform Engineer to lead the design, development, deployment, and maintenance of enterprise data and analytics platforms in Pune. You will architect scalable data pipelines, govern data quality, and drive automation across cloud-native lakehouse environments.

The role requires 8+ years in data engineering, strong leadership, and proficiency with Azure Databricks, Spark, Python, and Scala.

Qualifications

  • 8+ years of hands-on data engineering experience with strong leadership capabilities.
  • Experience designing and delivering enterprise data and analytics platforms.
  • Proven ability to partner with stakeholders and drive scalable data solutions.

Responsibilities

  • Design, develop, deploy, and maintain scalable data pipelines and analytics platforms.
  • Build reliable ETL/ELT pipelines and implement data governance, quality, and security.
  • Lead technical teams and mentor engineers in modern data engineering practices.
  • Collaborate with data scientists, analysts, and stakeholders to deliver data solutions.

Education

Bachelor's degree in a relevant technical discipline

Tools

Azure Databricks
Apache Spark
Python
Scala

Job description

to apply - email only, no card. You can also save this posting or score it againstyour profile with AI.## About the roleLeads the design, development, and maintenance of scalable data and analytics platforms while collaborating with stakeholders to deliver enterprise solutions. Implements data pipelines, governance practices, and automation to ensure high-quality, secure, and efficient data processing.## RequirementsRequires 8+ years of experience in data engineering with advanced proficiency in Azure Databricks, Spark, Python, and Scala. Candidates must possess strong technical leadership skills and a bachelor's degree in a relevant technical discipline.## Full descriptionLeads the design, development, deployment, and maintenance of data and analytics platforms. Develops reliable, scalable, and efficient data pipelines and data processing solutions that enable data to be effectively processed, stored, governed, and made available to analysts and other data consumers. Collaborates with business stakeholders, IT experts, data scientists, architects, and subject-matter experts to deliver enterprise data and analytics solutions aligned with business and technical requirements.Key Responsibilities* Design, develop, and automate distributed data ingestion and transformation solutions using data from relational, event-based, semi-structured, and unstructured sources.* Build reliable, scalable, and efficient ETL/ELT data pipelines using appropriate tools, technologies, and scripting languages.* Design and implement data quality, validation, monitoring, and alerting frameworks to identify and resolve data integrity issues.* Implement data governance practices covering metadata, data access, retention, compliance, and security.* Design and implement physical data models, including database structures, indexing, and table relationships, to support performance and scalability.* Develop and operate large-scale data storage and processing solutions across cloud and distributed data platforms, including data lakes, warehouses, and lakehouse environments.* Optimize data pipelines, Spark workloads, databases, and cloud infrastructure for performance, reliability, scalability, and cost efficiency.* Integrate data from a variety of enterprise applications and source systems and support real-time and event-driven data processing.* Develop automation for common and repeatable data preparation, integration, deployment, and platform-management activities to minimize manual and error-prone processes.* Implement CI/CD and DevOps practices to support automated deployment, testing, and release management.* Participate in troubleshooting, testing, validation, and continuous improvement of data pipelines and platform solutions.* Ensure data platforms and solutions meet applicable quality, governance, security, compliance, and regulatory requirements.* Collaborate with data scientists, analysts, architects, IT teams, and business stakeholders to understand requirements and deliver effective data solutions.* Document data solutions, processes, designs, and technical information to support knowledge transfer and operational effectiveness.* Apply Agile development methodologies such as Scrum and Kanban to deliver data engineering initiatives.* Provide technical leadership and mentor less experienced team members, promoting engineering excellence and collaboration.ResponsibilitiesSkills* Strong technical leadership and decision-making skills, with the ability to lead large-scale data engineering initiatives from concept through production deployment.* Strong problem-solving and analytical skills, including the ability to diagnose complex data pipeline, platform, and performance issues.* Excellent communication and collaboration skills, with the ability to work effectively with technical and business stakeholders.* Ability to translate complex technical concepts and stakeholder requirements into actionable data solutions.* Strong customer focus and ability to develop solutions aligned with business objectives.* Ability to balance strategic architecture decisions with hands-on technical execution.* Strong project management, prioritization, and organizational skills in a fast-paced environment.* Strong understanding of data quality, governance, security, compliance, and data management principles.* Passion for continuous improvement, emerging technologies, and modern data engineering practices.* Ability to mentor, guide, and develop technical talent while fostering a collaborative engineering culture.Technical SkillsData Engineering & Processing* Advanced proficiency in Azure Databricks, Apache Spark, and distributed data processing frameworks.* Strong expertise in enterprise-scale ETL/ELT and data ingestion pipeline design and development.* Experience processing structured, semi-structured, streaming, and large-scale datasets.* Strong understanding of Big Data technologies and scalable cloud-native data architectures.Programming & Development* Advanced proficiency in Python and Scala for large-scale data processing and engineering solutions.* Strong SQL skills for querying, transformation, optimization, and analysis of large datasets.* Experience with scripting, automation, version control, testing, and build processes.Cloud & Platform Technologies* Strong experience with Azure data services, including: - Azure Databricks* Azure Data Lake Storage (ADLS)* Azure Blob Storage* Azure Synapse Analytics* Azure SQL Data Warehouse* Strong understanding of cloud architecture principles, scalability, reliability, and security best practices.Data Storage & Lakehouse* Experience with modern data storage and lakehouse technologies, including: - Delta Lake* Apache Iceberg* Parquet* ORC* Strong understanding of data lake, data warehouse, and lakehouse architectures.Streaming & Data Integration* Experience with Kafka or similar real-time streaming and event-processing technologies.* Experience integrating data from a wide variety of enterprise applications and source systems.* Familiarity with Qlik Replicate or similar data replication and ingestion technologies is preferred.DevOps & Engineering Excellence* Experience implementing CI/CD pipelines and automated deployment processes.* Strong knowledge of Git, Jenkins, testing methodologies, version control, and release management.* Experience establishing coding standards, technical documentation, and engineering governance practices.Data Quality, Governance & Security* Experience implementing data validation frameworks, monitoring solutions, and data quality controls.* Knowledge of cloud security, access management, data governance, compliance, and regulatory requirements.* Experience designing reliable, auditable, and observable data platforms.Experience* 8+ years of hands-on experience in Data Engineering, Data Platform Engineering, Big Data, or a related discipline.* 3+ years of experience leading technical teams, projects, or significant data engineering initiatives.* Proven experience building and supporting large-scale cloud-based data ingestion, transformation, and analytics platforms.* Experience developing end-to-end ETL/ELT solutions in Azure cloud environments.* Demonstrated experience delivering high-performance and scalable distributed data processing solutions using Spark and Databricks.* Experience optimizing data pipelines, Spark workloads, databases, and cloud infrastructure for performance, reliability, and cost efficiency.* Experience working with modern data lake, data warehouse, and lakehouse architectures.* Experience implementing DevOps practices, CI/CD pipelines, and automated deployment processes.* Experience working within Agile software development methodologies.* Experience partnering with data scientists, analysts, architects, and business stakeholders to deliver enterprise data solutions.* Experience
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Engineer - Senior
Data Engineer - Senior

JobCubby • Pune District

On-site
INR 3,000,000 - 5,400,000
Associate Data Engineer
Associate Data Engineer

JobCubby • Hyderabad

On-site
INR 1,200,000 - 2,000,000
Senior Databricks Engineer
Senior Databricks Engineer

Keka Technologies Private Limited • Hyderabad

On-site
INR 1,500,000 - 2,500,000
Sr. Lead- Databricks
Sr. Lead- Databricks

Azimuth Grc • Gurugram District

On-site
INR 4,000,000 - 7,000,000
Technical Lead - Data Engineer ( Databricks Azure)
Technical Lead - Data Engineer ( Databricks Azure)

Srijan: Now Material • Gurugram District

On-site
INR 2,400,000 - 4,800,000
Senior/Lead Data Engineer
Senior/Lead Data Engineer

ICICI Lombard • Mumbai

On-site
INR 2,800,000 - 4,000,000
Senior Databricks Engineer
Senior Databricks Engineer

DataBeat • Hyderabad

On-site
INR 3,500,000 - 5,500,000
Data Engineer Lead
Data Engineer Lead

Skillventory • Navi Mumbai, Mumbai

On-site
INR 900,000 - 1,500,000
Fullstack Lead- ITO Transition
Fullstack Lead- ITO Transition

Daimler AG • Bengaluru

Hybrid
INR 1,800,000 - 2,600,000
Subcon - Data Engineer
Subcon - Data Engineer

JobCubby • Pune District

On-site
INR 1,200,000 - 2,100,000