Note: This role is strictly for a Cloudera Architect with strong development and architecture experience (Spark, Hive, CDP, Data Engineering).
Experience: 12+ Years
Key Responsibilities
- Design and architect scalable Big Data solutions using Cloudera Data Platform (CDP).
- Lead end-to-end implementation of Hadoop-based data platforms.
- Design data lakes, lakehouse architectures, and real-time data processing systems.
- Work closely with business stakeholders to translate requirements into technical solutions.
- Define data governance, security, and compliance frameworks.
- Optimize performance tuning for Hive, Spark, and Impala workloads.
- Lead cluster planning, sizing, capacity management, and DR strategy.
- Provide technical leadership and mentor data engineers.
- Conduct architecture reviews and ensure best practices are followed.
Required Technical Skills
- Strong experience with Cloudera Data Platform (CDP Private Cloud / Public Cloud)
- Experience in cluster management using Cloudera Manager
- Knowledge of Ranger & Atlas for security and governance
- Strong experience in Spark (Scala/PySpark)
- Batch and real-time processing architecture
- ETL/ELT design and optimization
Cloud Experience (Preferred)
- AWS / Azure / GCP integration with CDP
- Experience with object storage (S3/ADLS/GCS)
Soft Skills
- Strong stakeholder communication
- Leadership and mentoring skills
- Excellent problem-solving ability
- Experience handling enterprise clients
Good to Have
- Experience with Delta Lake / Iceberg
- Knowledge of Data Governance frameworks
- Exposure to ML/AI pipelines
Please share resume on Dhanashree.C@asplinfo.com