Big Data Platform Engineer (Relocation to Malaysia)

Virtej Technologies

Karachi Division

On-site

PKR 2,500,000 - 4,500,000

Full time

4 days ago
Be an early applicant
Application generator

Stand out for this role — generate a tailored resume and cover letter in about a minute.

Get past ATS filters

Job summary

Virtej Technologies seeks a highly motivated Big Data Platform Engineer to design, develop, and maintain scalable data platforms and pipelines. The role emphasizes Python, Spark (PySpark), SQL, and Linux-based environments with DevOps practices.

Responsibilities include building robust data pipelines, ETL/ELT processing, and optimizing SQL and data models. You will work with cross-functional teams to ensure platform reliability and performance across enterprise-scale analytics.

Qualifications

  • 6 to 12 years (or 8 to 12+ years) in Data Engineering or Big Data Platform Development.
  • Experience with enterprise-scale data platforms.
  • Experience in distributed computing environments.
  • Strong Python/Spark/SQL proficiency.
  • Linux and DevOps practices exposure.

Responsibilities

  • Design, develop, and maintain scalable data pipelines using Python and Apache Spark.
  • Implement efficient ETL/ELT processes for large-scale datasets.
  • Develop and optimize complex SQL queries, data models, and transformations.
  • Ensure data quality, integrity, and reliability across the platform.

Skills

Python
Apache Spark (PySpark)
SQL
Apache Airflow
Kubernetes AKS/SKE
MinIO / S3 storage
ETL/ELT Processing
Performance Tuning
Git Version Control
Agile/Scrum
Shell Scripting (Bash/KSH)
CI/CD concepts
Cloud platforms (Azure/AWS)

Tools

Docker
AWS
Azure
GKE/EKS
Linux

Job description

We are seeking a highly motivated and skilled Big Data Platform Engineer to design, develop, and maintain scalable data processing platforms and pipelines. The ideal candidate should have strong expertise in Python, Apache Spark, SQL, and experience working in Linux-based environments with exposure to DevOps practices. The role involves building reliable, high-performance data solutions that support enterprise-scale analytics and business-critical applications.

Key Responsibilities

  • Design, develop, and maintain scalable data pipelines using Python and Apache Spark.
  • Implement efficient ETL/ELT processes for large-scale structured and unstructured datasets.
  • Develop and optimize complex SQL queries, data models, and transformations.
  • Ensure data quality, integrity, and reliability across the platform.

Platform Operations

  • Work with Linux-based environments for deployment, troubleshooting, and performance tuning.
  • Develop and maintain shell scripts for automation and operational tasks.
  • Monitor and optimize Spark jobs for performance, scalability, and resource utilization.

DevOps & Automation

  • Implement CI/CD pipelines and deployment automation.
  • Participate in infrastructure provisioning, monitoring, and release management activities.
  • Collaborate with DevOps teams to improve platform reliability and operational efficiency.
  • Work closely with Data Architects, Product Owners, and Business Stakeholders.
  • Participate in code reviews and ensure adherence to engineering best practices.
  • Create and maintain technical documentation and operational runbooks.

Mandatory Skills

Core Big Data Platform Skills

  • Strong programming experience in Python
  • Hands-on expertise with Apache Spark (PySpark preferred)
  • Strong SQL development and query optimization skills
  • Apache Airflow for workflow orchestration and scheduling
  • Kubernetes (AKS/SKE) for container orchestration and deployment. AKS/EKS/GKE also fine.
  • MinIO / S3 Compatible Object Storage like AWS S3 or other S3-compatible object storage experience
  • ETL/ELT Processing
  • Performance Tuning
  • Version Control (Git)
  • Agile/Scrum Delivery Model

Good to Have Skills

  • Shell Scripting (Bash/KSH)
  • Understanding of DevOps practices and CI/CD pipelines
  • Docker and Containerization Concepts
  • Cloud Platform Experience (Azure/AWS)

Desired Experience

  • 6 to 8 years of experience or 8 to 12 years or 12+ years of experience in Data Engineering or Big Data Platform Development .
  • Experience working with enterprise-scale data platforms.
  • Experience in distributed computing environments.

Strong analytical and problem-solving skills.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Data Engineer
Data Engineer

Zorba Consulting • Hyderabad City Taluka

On-site
INR 1,200,000 - 2,400,000
Platform Engineer( On-Premises)
Platform Engineer( On-Premises)

Datamatics Technologies LLC • Karachi Division

On-site
PKR 13,284,000 - 17,712,000
Senior Databricks Engineer
Senior Databricks Engineer

Digifloat • Islamabad

On-site
PKR 2,600,000 - 3,400,000
Senior Software Engineer - Data Engineering & AI
Senior Software Engineer - Data Engineering & AI

Devsinc, LLC • Islamabad

On-site
PKR 1,674,000 - 3,348,000
Data Architect - Microsoft Azure Data Services, DataLake, Databricks
Data Architect - Microsoft Azure Data Services, DataLake, Databricks

HireOn • Pakistan

Hybrid
PKR 3,000,000 - 5,500,000
Senior Manager Data
Senior Manager Data

Techsurge Private Limited • Karachi Division

On-site
PKR 3,000,000 - 6,000,000
Senior Software Engineer - Data Engineering & AI
Senior Software Engineer - Data Engineering & AI

Devsinc 17 • Islamabad

On-site
PKR 1,800,000 - 2,800,000
Senior Data Engineer – Cloud Platforms (Azure | GCP)
Senior Data Engineer – Cloud Platforms (Azure | GCP)

Tkxel LLC • Lahore

On-site
PKR 33,213,000 - 49,820,000
Senior Data Engineer – Cloud Platforms (Azure | GCP)
Senior Data Engineer – Cloud Platforms (Azure | GCP)

Zohorecruit • Lahore

On-site
PKR 30,446,000 - 44,285,000
Senior Data Engineer
Senior Data Engineer

NorthBay Solutions LLC • Karachi Division

On-site
PKR 1,500,000 - 2,500,000