Get more replies from employers
Send a job-specific resume in minutes.
E4 Software Services Pvt Ltd. is seeking a senior data engineer with 6+ years of experience to design and maintain scalable data pipelines using PySpark and AWS Databricks. You will transform structured and unstructured data, implement production-grade code, and ensure best practices for data quality, security, and governance.
Expertise in PySpark, Spark clusters, SQL, AWS S3, and cloud-native architectures is required, along with CI/CD for data pipelines and performance tuning.
· 6+ years of professional and relevant experience in software industry.
·Strong hands-on expertise in PySpark for distributed data processing and transformation.
· Proven experience with Python programming, including implementing reusable, production-grade code.
· Practical knowledge of AWS Databricks for building and orchestrating large-scale data pipelines.
· Demonstrated experience in processing structured and unstructured data using Spark clusters and cloud data platforms.
· Ability to apply data engineering best practices including version control, CI/CD for data pipelines, and performance tuning.
Preferred Skills:
· Working knowledge of SQL for data querying, analysis, and troubleshooting.
· Experience using AWS S3 for object storage and EC2 for compute orchestration in cloud environments.
· Understanding of cloud-native data architectures and principles of data security and governance.
· Exposure to BFSI domain use cases and familiarity with handling sensitive financial data.
· Familiarity with CI/CD tools and cloud-native deployment practices.
· Familiarity with ETL scripting and data pipeline automation.
· Knowledge of big data ecosystems and distributed computing.