Senior Data Engineer

Crisil

Maharashtra

On-site

INR 1,800,000 - 3,000,000

Full time

4 hours ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

Crisil seeks a skilled data architect to implement scalable data architecture across on-prem and cloud platforms (Azure/GCP/AWS). You will work with data engineers and product teams to design data models and pipelines using Spark, Kafka, and Delta Lake.

The role emphasizes data quality, security, and performance in data warehouses and marts. The ideal candidate will design, optimize, and maintain batch and streaming workflows, leveraging Databricks and related tools to support BI and analytics

Qualifications

  • Strong data analysis and modeling skills.
  • Experience designing scalable data architectures.
  • Proficiency in SQL, query optimization, and database design.
  • Hands-on experience with Spark, Databricks, and data pipelines.
  • familiarity with Delta Lake, data warehousing, and BI ingestion.

Responsibilities

  • Implement scalable data architecture on on-prem and cloud platforms.
  • Collaborate with data engineers and product teams to model data needs.
  • Develop batch and streaming pipelines using Spark, Databricks, Kafka, and Flume.
  • Integrate data from relational, NoSQL, APIs, and files.
  • Design Delta Lake and data warehouses for BI and analytics.
  • Maintain data pipelines and ensure data quality and performance.
  • Optimize queries, storage, and data processing workflows.
  • Implement data security, access controls, and backups.

Skills

Data analysis
Data modeling
SQL knowledge
Data visualization
Requirements gathering

Tools

Apache Spark
PySpark
Kafka
Flume
Delta Lake
Databricks
Azure
GCP
AWS
LangChain
OpenAI API
Vertex AI
Amazon Bedrock
Kinesis
BigQuery

Job description

  • Implement Data Architecture: Implement scalable, secure, and efficient data architecture on on-prem and cloud platforms (Azure/GCP/AWS) to support business growth and data-driven decision-making.
  • Collaborate with data engineers, and product teams to identify data requirements and develop data models that meet business needs.
  • Data Ingestion and Integration:
  • Develop and maintain data ingestion pipelines using various tools and technologies, such as Apache Spark, PySpark, Kafka, and Flume.
  • Experience with GenAI tooling: LangChain, OpenAI API, Amazon Bedrock, or Vertex AI integrations, Claude, Copilot
  • Integrate data from multiple sources, including relational databases, NoSQL databases, APIs, and files.
  • Batch and Stream Processing:
  • Develop and maintain batch and stream processing pipelines using tools like Apache Spark & databricks.
  • Integrate with messaging systems, such as Apache Kafka, Amazon Kinesis, and Google Cloud Pub/Sub.
  • SQL Knowledge:
  • Very strong SQL knowledge, including query optimization, indexing, and database design.
  • Delta Lake and Data Warehouse:
  • Design and implement Delta Lake and data warehouse/mart solutions to support business intelligence, reporting, and analytics.
  • Develop and maintain data pipelines to ingest, process, and store data in Delta Lake and data warehouses.
  • Distributed Databases and Data Warehousing:
  • Implement and maintain data warehouses, such as Amazon Redshift, Google BigQuery, and Azure Synapse Analytics.
  • Database Design and Development:
  • Design, develop, and maintain efficient and scalable database systems across different platforms of relational databases (such as Oracle, MySQL, PostgreSQL, SQL Server).
  • Collaborate with cross-functional teams to understand data requirements and translate them into effective database solutions.
  • Implement and design data models and database schemas that align with business needs, ensuring data integrity and efficient data retrieval.
  • Develop and optimize database queries, stored procedures, and functions for maximum performance and responsiveness.
  • Performance Tuning and Optimization:
  • Analyze and monitor database performance using diagnostic tools, identifying and resolving performance bottlenecks and inefficiencies.
  • Optimize data processing workflows and queries to improve performance, reduce latency, and increase throughput.
  • Data Management: Implement data archival mechanisms and data retention policies to ensure efficient data storage.
  • Ensure the security and integrity of data by implementing access controls, data encryption, and backup and recovery strategies.
  • Automation and Integration:
  • Identify and implement automation solutions for data workflows.
  • Collaborate with the development team to integrate database solutions into software applications effectively.
  • Data Mart and Data Lake: Design and implement data marts and data lakes to support business intelligence, reporting, and analytics.
  • Develop and maintain data pipelines to ingest, process, and store data in data lakes, such as Apache Hadoop, Amazon S3, and Azure Data Lake Storage.
  • CI/CD and Automation:
  • Develop and maintain automated testing, deployment, and monitoring scripts using tools like Jenkins, GitLab CI/CD, or similar.
  • Ensure continuous integration and delivery of data pipelines and applications.
  • Data Analysis and Modeling:
  • Strong data analysis skills, including data modeling, data mining, and data visualization.
  • Collaborate with data modelers to develop and implement data models to drive business insights and decision-making.
  • Analyze complex data sets to identify trends, patterns, and correlations.
  • Exploration of New Tools:
  • Ability to explore new tools and technologies, and quickly develop proof-of-concepts (POCs) for data engineering open-source tools.
  • Documentation: Document database design, configurations, and technical specifications.
Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

Crisil • Mumbai, Pune District, Hyderabad

On-site
INR 1,800,000 - 2,600,000
Data Architect
Data Architect

Laksh Human Resource • Dadri

On-site
INR 1,400,000 - 2,100,000
Data Architect
Data Architect

PeopleStrong • India

On-site
INR 2,500,000 - 5,000,000
Senior Data Engineer
Senior Data Engineer

PeopleStrong • India

On-site
INR 900,000 - 1,500,000
Data Architect
Data Architect

Pashtek • Salesforce Partner | Data & AI • India

Hybrid
INR 3,000,000 - 6,000,000
Data Architect
Data Architect

New Era India • Bengaluru

Hybrid
INR 4,000,000 - 7,000,000
Data Engineer
Data Engineer

fluid.live • Chennai District

On-site
INR 1,200,000 - 1,800,000
Senior Data Engineer
Senior Data Engineer

Proclink • Gandhamguda

On-site
INR 800,000 - 1,500,000
Technical Lead - Data Engineer (Data&AI)
Technical Lead - Data Engineer (Data&AI)

Srijan Technologies PVT LTD • Gurugram District

On-site
INR 4,000,000 - 7,500,000
Solution Architect
Solution Architect

Cyient • Hyderabad

On-site
INR 900,000 - 1,400,000
Microsoft Azure Data Engineer–certs
Azure Solutions Architect Expert
AWS Certified Data Analytics - Special
+2