Data Engineering, Data Warehousing, and Cloud-based Data Platforms Job Summary Lead Engineer - II with deep expertise in Snowflake and Matillion ETL to design, build, and optimize scalable data pipelines and data warehouse solutions. The candidate should have strong hands-on experience in cloud platforms, data integration, performance tuning, and end-to-end data lifecycle management. ________________________________________
Key Responsibilities
- Design and implement scalable data solutions using Snowflake
- Develop, maintain, and optimize ETL/ELT pipelines using Matillion
- Build and manage data ingestion frameworks for structured and semi-structured data
- Collaborate with business stakeholders to understand data requirements and translate them into technical solutions
- Optimize Snowflake performance (query tuning, clustering, warehouse sizing, etc.)
- Implement data modeling techniques (Star Schema, Snowflake Schema, Dimensional Modeling)
- Ensure data quality, governance, and security compliance
- Integrate data from multiple sources such as APIs, databases, flat files, and external systems
- Work with cloud platforms (AWS/Azure/GCP) for end-to-end data pipeline deployment
- Mentor junior engineers and lead technical discussions
- Automate workflows and monitoring using best practices
- Participate in architecture design and code reviews
Required Skills
Core Skills
- Strong hands-on experience with Snowflake Data Warehouse
- Snowflake architecture, Time Travel, Data Sharing, Zero Copy Cloning
- Performance tuning and cost optimization
- Expertise in Matillion ETL/ELT
- Job orchestration, transformations, and scheduling
- Advanced knowledge of SQL & query optimization
- Data Engineering
- Experience with ETL/ELT tools and data pipeline development
- Strong understanding of data warehousing concepts
- Experience with semi-structured data (JSON, XML, Parquet)
Cloud & Integration
- Hands-on experience with at least one cloud platform (AWS/Azure/GCP)
- Experience with services like:
- AWS: S3, Lambda, Redshift (optional), Glue
- Azure: ADLS, Data Factory
- API integration and data ingestion techniques
Programming
- Proficiency in Python / Shell scripting for automation
- Familiarity with version control tools (Git)
Preferred Skills
- Knowledge of dbt (Data Build Tool)
- Experience with Airflow or other orchestration tools
- Familiarity with CI/CD pipelines and DevOps practices
- Experience in data governance and security frameworks
- Domain experience in healthcare, finance, or retail (optional)
Nice to Have
- Snowflake certification (SnowPro Core/Advanced)
- Matillion certification
- Experience in handling large-scale distributed data systems