Job Location: Lincoln Rhode Island (Onsite)
Primary skills: Snowflake, PySpark
Secondary skills: SQL, Python
Mode of Work: Work from Office
About the job
We are seeking a skilled Snowflake Engineer with expertise in PySpark to join our team and contribute to building scalable, efficient, and governed data solutions. The ideal candidate will have hands-on experience in designing and optimizing Snowflake data warehouses, leveraging PySpark for complex ETL/ELT processes. Proficiency in advanced Snowflake features, such as streams, tasks, materialized views, and SQL, is essential. A solid understanding of data governance principles, metadata management, and cloud platforms like AWS or Azure will be highly beneficial.
Responsibilities
- Design, develop, and manage data pipelines in Snowflake to support scalable and efficient data warehousing solutions.
- Implement and maintain PySpark notebooks, ensuring seamless data ingestion, transformation, and loading into Snowflake.
- Develop and optimize PySpark scripts for advanced data transformation and workflow automation.
- Integrate AWS with Snowflake stages for enhanced data processing
- Collaborate with cross-functional teams to gather requirements and translate business needs into robust technical solutions.
- Monitor and troubleshoot Snowflake, and Pyspark workflows to ensure performance, reliability, and availability.
- Implement and optimize data partitioning, indexing, and caching techniques to enhance query performance in Snowflake.
- Automate repetitive tasks and implement CI/CD pipelines for seamless deployment and monitoring of data workflows.
- Collaborate with stakeholders to gather requirements and translate them into technical solutions.
- Build and maintain comprehensive mapping documents, transformation of business rules, and custom processes for data governance and analytics.
Requirements - Must Have
- 5-7 years of experience in data engineering with expertise in Snowflake, and PySpark.
- Strong hands-on experience with Snowflake, including designing schemas, managing pipelines, and optimizing query performance.
- Proficiency in developing and maintaining Snowpark workflows, with expertise in integrating AWS with Snowflake.
- In-depth knowledge of Snowflake for implementing data governance, metadata management, and data lineage processes.
- Advanced skills in PySpark for complex data transformation and automation tasks
- Solid understanding of Data Warehousing concepts and distributed computing.
- Experience with Agile/Scrum methodologies and excellent problem-solving capabilities.
- Effective verbal and written communication skills for direct client interaction.
- Experience working as both an individual contributor and a team player.
- Experience in the P&C Insurance domain is preferred.