Senior PySpark Data Architect: Scalable Pipelines & ETL

iSystems Ltd

Punjab

On-site

PKR 3,500,000 - 5,500,000

Full time

5 days ago
Be an early applicant
Application generator

An application made for this job — a tailored resume and cover letter that speak straight to the posting.

Get past ATS filters

Job summary

iSystems Ltd is seeking a Data Architect with strong hands-on PySpark experience to design, develop, optimize, and maintain scalable data processing solutions. You will work with large-scale data processing and distributed computing environments to deliver production-grade PySpark code.

The role requires 8+ years in Data Engineering, expertise in PySpark, Spark architecture, DataFrames, and cloud/on-premise data pipelines.

Qualifications

  • 8+ years of experience in Data Engineering or related field.
  • Strong and mandatory hands-on experience with PySpark.
  • Proven experience writing production-grade PySpark code.
  • Strong understanding of Spark architecture, DataFrames, Spark SQL, transformations, actions, partitioning, caching, and optimization.
  • Strong proficiency in Python with practical data engineering application development.
  • Hands-on experience building and managing large-scale ETL/ELT data pipelines.
  • Strong SQL skills with complex queries, joins, aggregations, and performance optimization.
  • Experience with cloud data platforms such as Azure, AWS, or GCP is preferred.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using PySpark.
  • Write clean, efficient, and production-ready PySpark code for large-scale data processing.
  • Develop ETL/ELT pipelines to ingest, transform, validate, and load data from multiple sources.
  • Perform complex data transformations, aggregations, joins, filtering, and data cleansing using PySpark.
  • Optimize PySpark jobs for performance, scalability, memory utilization, and execution time.
  • Work with distributed data processing frameworks and large datasets in cloud or on-premise environments.
  • Develop reusable PySpark frameworks, libraries, and data processing components.
  • Troubleshoot and resolve data pipeline failures, performance bottlenecks, and data quality issues.
  • Implement data validation, error handling, logging, and monitoring within data pipelines.
  • Work closely with Data Architects, Data Engineers, Data Scientists, BI teams, and business stakeholders to understand data requirements.
  • Participate in data modeling, pipeline architecture, and technical design discussions.
  • Review code and provide technical guidance and mentorship to junior and mid-level Data Engineers.
  • Ensure adherence to coding standards, data engineering best practices, security, and governance requirements.
  • Work with CI/CD processes and version control systems for deploying and managing data pipelines.
  • Contribute to technical documentation and maintain clear documentation of data pipelines and processes.

Skills

PySpark
Spark Architecture
Python
SQL
ETL/ELT
CI/CD
Distributed Computing
Databricks
Delta Lake
Kafka
Hive
Git

Tools

Databricks
Delta Lake
Kafka
Hive
Git

Job description

iSystems Ltd is seeking a Data Architect with strong hands-on PySpark experience to design, develop, optimize, and maintain scalable data processing solutions. You will work with large-scale data processing and distributed computing environments to deliver production-grade PySpark code.

The role requires 8+ years in Data Engineering, expertise in PySpark, Spark architecture, DataFrames, and cloud/on-premise data pipelines.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior PySpark Data Architect — Scalable Pipelines
Senior PySpark Data Architect — Scalable Pipelines

iSystems Ltd • Punjab

On-site
PKR 1,200,000 - 1,800,000
Data Architect - Pyspark
Data Architect - Pyspark

iSystems Ltd • Punjab

On-site
PKR 3,500,000 - 5,500,000
Senior Data Engineer: PySpark & Databricks
Senior Data Engineer: PySpark & Databricks

Strategic Systems International • Lahore

On-site
PKR 1,800,000 - 3,200,000
Senior Data Engineer (Pyspark, Databricks)
Senior Data Engineer (Pyspark, Databricks)

Strategic Systems International • Lahore

On-site
PKR 1,800,000 - 3,200,000
Senior Data Engineer — Real-Time Data Platform Lead
Senior Data Engineer — Real-Time Data Platform Lead

Systems Limited • Punjab

On-site
PKR 2,400,000 - 3,600,000
Senior Data Engineer: Scalable Pipelines & Cloud Data
Senior Data Engineer: Scalable Pipelines & Cloud Data

Abacus • Lahore

On-site
PKR 240,000 - 480,000
Competitive pay
Large-scale projects
Growth environment
+2
Senior Data Engineer
Senior Data Engineer

Systems Limited • Punjab

On-site
PKR 2,400,000 - 3,600,000
Associate Data Engineer Team Lead: Build Pipelines & Impact
Associate Data Engineer Team Lead: Build Pipelines & Impact

Devsinc • Lahore

On-site
Confidential
Data Engineer
Data Engineer

Zorba Consulting • Hyderabad City Taluka

On-site
INR 1,200,000 - 2,400,000
Data Architect - Microsoft Azure Data Services, DataLake, Databricks
Data Architect - Microsoft Azure Data Services, DataLake, Databricks

HireOn • Pakistan

Hybrid
PKR 3,000,000 - 5,500,000