Spark Data Engineer

Persistent Systems

Pune District

Hybrid

INR 2,500,000 - 4,500,000

Full time

5 days ago
Be an early applicant

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Hybrid work model
Flexible work hours
Company-sponsored higher education & |

Job summary

Persistent Systems is seeking a Spark Data Engineer to design, develop, and optimize large-scale data pipelines using Apache Spark, with Java/Scala development, data engineering best practices, and CI/CD for data platforms. The role emphasizes building batch and real‑time workflows and improving performance in cloud environments.

You will collaborate across data architects, analysts, and cross‑functional teams to define requirements, implement standards, and ensure high‑quality, maintainable

Qualifications

  • Hands-on Spark development with real projects and production workloads.
  • Experience designing scalable data pipelines using Spark, Java and Scala.
  • Exposure to CI/CD for data platforms and cloud deployment.

Responsibilities

  • Design, develop, and optimize large-scale data processing pipelines using Apache Spark.
  • Build and maintain batch and real-time data ingestion and transformation workflows.
  • Develop scalable data solutions using Java and/or Scala.
  • Analyze and improve Spark job performance, resource utilization, and execution efficiency.
  • Collaborate with data architects, business analysts, and cross-functional teams to understand requirements and deliver data solutions.
  • Implement coding standards, unit testing, and best practices for data engineering projects.
  • Build and maintain CI/CD pipelines for automated deployment and testing of data applications.
  • Troubleshoot production issues and provide timely resolution.

Skills

Apache Spark
Java
Scala
CI/CD
Hadoop
Cloud platforms
ETL/ELT
Data warehousing

Tools

Jenkins
GitHub Actions
GitLab CI
Azure DevOps

Job description

We are seeking a highly skilled Spark Data Engineer with strong hands‑on experience in building scalable data processing solutions using Apache Spark. The ideal candidate should have experience in Java/Scala development, data engineering best practices, and CI/CD implementation for data platforms.

  • Location: All Persistent Locations
  • Experience: 7 to 13 Years
  • Job Type: Full‑Time Employment
What You’ll Do:
  • Design, develop, and optimize large‑scale data processing pipelines using Apache Spark.
  • Build and maintain batch and real‑time data ingestion and transformation workflows.
  • Develop scalable data solutions using Java and/or Scala.
  • Analyze and improve Spark job performance, resource utilization, and execution efficiency.
  • Collaborate with data architects, business analysts, and cross‑functional teams to understand requirements and deliver data solutions.
  • Implement coding standards, unit testing, and best practices for data engineering projects.
  • Build and maintain CI/CD pipelines for automated deployment and testing of data applications.
  • Troubleshoot production issues and provide timely resolution.
  • Participate in code reviews and ensure high‑quality, maintainable code.
  • Support cloud‑based and distributed data processing environments.
Expertise You’ll Bring:
  • Apache Spark (Core, Spark SQL, Data Frames, Performance Tuning)
  • Java (Intermediate to Advanced)
  • Scala (Basic to Intermediate)
  • CI/CD tools such as Jenkins, GitHub Actions, GitLab CI, Azure DevOps, etc.
  • Experience with Hadoop ecosystem (HDFS, Hive, YARN).
  • Experience working on data migration or modernization projects.
  • Knowledge of cloud platforms such as AWS, Azure, or GCP.
  • Understanding of data warehousing and ETL/ELT concepts.
  • Strong problem‑solving and analytical skills.
  • Look for candidates who can demonstrate:
  • Hands‑on Spark development experience (not just listed as a skill).
  • Performance tuning and optimization of Spark jobs.
  • Real project experience using Java/Scala with Spark.
  • End‑to‑end data pipeline development and deployment experience.
  • Competitive salary and benefits package
  • Culture focused on talent development with quarterly growth opportunities and company‑sponsored higher education and certifications
  • Opportunity to work with cutting‑edge technologies
  • Employee engagement initiatives such as project parties, flexible work hours, and Long Service awards
  • Insurance coverage: group term life, personal accident, and Mediclaim hospitalization for self, spouse, two children, and parents
Values‑Driven, People‑Centric & Inclusive Work Environment:

Persistent is dedicated to fostering diversity and inclusion in the workplace. We invite applications from all qualified individuals, including those with disabilities, and regardless of gender or gender preference. We welcome diverse candidates from all backgrounds.

  • We support hybrid work and flexible hours to fit diverse lifestyles.
  • Our office is accessibility‑friendly, with ergonomic setups and assistive technologies to support employees with physical disabilities.
  • If you are a person with disabilities and have specific requirements, please inform us during the application process or at any time during your employment

“Persistent is an Equal Opportunity Employer and prohibits discrimination and harassment of any kind.”

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Azure Cloud Data Engineer
Azure Cloud Data Engineer

Persistent Systems • Pune District

Hybrid
INR 2,000,000 - 4,000,000
Education sponsorship
Long service awards
Insurance coverage
Data Engineer
Data Engineer

Persistent Systems • Pune District

Hybrid
INR 1,000,000 - 1,500,000
Competitive salary
Quarterly growth opportunities
Company-sponsored education and certifications
+2
Azure Cloud Data Engineer
Azure Cloud Data Engineer

United States Digital Space LLC • Maharashtra

Hybrid
INR 1,800,000 - 2,500,000
Competitive salary
Hybrid work options
Education & certification sponsorship
Senior AWS Data Engineer
Senior AWS Data Engineer

Persistent Systems • Pune District

Hybrid
INR 4,000,000 - 7,000,000
Hybrid work
Education sponsorship
Flexible hours
+2
Sr. Azure Databricks Engineer
Sr. Azure Databricks Engineer

Persistent Systems • Bengaluru

Hybrid
INR 2,800,000 - 4,500,000
Senior Azure Data Engineer
Senior Azure Data Engineer

Persistent Systems • Mumbai

Hybrid
INR <1,000
Hybrid work
Insurance coverage
Education sponsorship
+3
Software Engineer
Software Engineer

Alegeus • Bengaluru

On-site
INR 3,500,000 - 5,500,000
AWS Architect
AWS Architect

Persistent Systems • Pune District

Hybrid
INR 4,000,000 - 7,000,000
Competitive salary
Hybrid work model
Education sponsorship and certs
+2
AWS Architect
AWS Architect

Persistent • Pune District

Hybrid
INR 4,000,000 - 6,000,000
Competitive salary
Hybrid work
Education sponsorship
+3
Data Engineer
Data Engineer

Persistent Systems • Bengaluru

Hybrid
INR 900,000 - 1,800,000
Hybrid work model
Professional development support
Medical insurance
+2