A leading software solutions company in Mumbai is seeking a Data Engineer to design and deploy high-performance data solutions. The ideal candidate will have strong experience with Python, PySpark, and SQL, as well as proficiency in AWS services, particularly with Databricks. This role requires collaboration with cross-functional teams and a commitment to best coding practices. Competitive salary offered.
Qualifications
4 to 8 years of experience in Databricks and big data frameworks.
Proficient in AWS services and data migration.
Experience in Unity Catalogue.
Familiarity with batch and real-time processing.
Responsibilities
Design, develop, test, and deploy data solutions using Python, PySpark, and SQL.
Collaborate with teams to understand requirements.
Implement efficient and maintainable code.
Implement and optimize data pipelines using AWS services.
Manage relational databases and write complex SQL queries.
Skills
Python
PySpark
SQL
AWS services
Databricks
Data engineering
Data Pipeline Optimization
Education
Bachelor’s degree in Computer Science or related field
Tools
AWS
Databricks
Job description
Responsibilities
Design, develop, test, and deploy high-performance and scalable data solutions using Python, PySpark, SQL
Collaborate with cross-functional teams to understand business requirements and translate them into technical specifications.
Implement efficient and maintainable code using best practices and coding standards.
AWS & Databricks Implementation:
Work with Databricks platform for big data processing and analytics.
Develop and maintain ETL processes using Databricks notebooks.
Implement and optimize data pipelines for data transformation and integration.
Utilize AWS services (e.g., S3, Glue, Redshift, Lambda) and Databricks to build and optimize data migration pipelines.
Leverage PySpark for large-scale data processing and transformation tasks.
Stay updated on the latest industry trends, tools, and technologies related to Python, SQL, and Databricks.
Share knowledge with the team and contribute to a culture of continuous improvement.
SQL Database Management:
Utilize expertise in SQL to design, optimize, and maintain relational databases.
Write complex SQL queries for data retrieval, manipulation, and analysis
Qualifications & Skills
Education: Bachelor’s degree in Computer Science, Engineering, Data Science, or a related field. Advanced degrees are a plus.
4 to 8 Years of experience in Databricks and big data frameworks
Proficient in AWS services and data migration
Experience in Unity Catalogue
Familiarity with Batch and real time processing
Data engineering with strong skills in Python, PySpark, SQL
Certifications: AWS Certified Solutions Architect, Databricks Certified Professional, or similar are a plus.
Soft Skills
Strong problem-solving and analytical skills.
Excellent communication and collaboration abilities.
Ability to work in a fast-paced, agile environment