Senior Java Spark Data Platform Architect

Veriipro

Berkeley Heights (NJ)

On-site

USD 140,000 - 190,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Veriipro in Berkeley Heights, NJ, seeks a senior data engineer to architect and implement scalable data pipelines using Spark (Java). You will lead batch and streaming ETL/ELT, optimize performance, and set coding standards across the team.

You will mentor engineers, drive data architecture decisions, partner with product and analytics, own production reliability, and help control cost and capacity for cluster infrastructure.

Qualifications

  • Bachelor’s or Master’s degree in Computer Science, Engineering, or related field.
  • 7+ years of professional Java development experience.
  • 5+ years hands-on experience with Apache Spark in production environments.
  • Expert-level understanding of distributed systems: fault tolerance, data locality, shuffle mechanics, resource management.
  • Proven track record designing systems processing terabyte+ scale data.
  • Strong SQL skills and deep familiarity with columnar storage formats (Parquet, ORC, Avro, Delta Lake/Iceberg).
  • Experience with cluster managers (YARN, Kubernetes) and cloud-managed Spark.
  • Proficiency with Kafka
  • Strong grasp of CI/CD, containerization, and infrastructure-as-code practices.

Responsibilities

  • Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java)
  • Lead design of batch and streaming ETL/ELT systems handling large data volumes
  • Deep-dive performance tuning: partitioning strategy, memory management, shuffle/skew optimization, job cost reduction
  • Set coding standards and lead code/design reviews across the team
  • Drive technical decisions on data architecture, storage formats, and pipeline orchestration
  • Mentor mid-level and junior engineers; act as a technical escalation point
  • Partner with product, analytics, and platform teams to translate requirements into scalable systems
  • Own production reliability — on-call ownership, incident response, root‑cause analysis for pipeline failures
  • Evaluate and introduce new tools/frameworks where they improve the system
  • Contribute to capacity planning and cost optimization for cluster infrastructure

Skills

Java
Spark
Distributed systems
SQL
Kafka
CI/CD
Containers
Kubernetes
YARN
Delta Lake

Education

Bachelor’s or Master’s degree in CS/Engineering

Tools

YARN
Kubernetes
Docker

Job description

Veriipro in Berkeley Heights, NJ, seeks a senior data engineer to architect and implement scalable data pipelines using Spark (Java). You will lead batch and streaming ETL/ELT, optimize performance, and set coding standards across the team.

You will mentor engineers, drive data architecture decisions, partner with product and analytics, own production reliability, and help control cost and capacity for cluster infrastructure.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Big Data Engineer - Scalable Pipelines & Spark
Big Data Engineer - Scalable Pipelines & Spark

Veriipro • Jersey City (NJ)

On-site
USD 100,000 - 130,000
Senior Big Data Lead — Spark, Hadoop & Cloud
Senior Big Data Lead — Spark, Hadoop & Cloud

Veriipro • United States

On-site
USD 180,000 - 240,000
Senior Java Spark Engineer - Scalable Data Processing
Senior Java Spark Engineer - Scalable Data Processing

Tech Mirrors • Berkeley Heights (NJ)

Hybrid
USD 120,000 - 150,000
Senior PySpark & Java Data Pipeline Engineer
Senior PySpark & Java Data Pipeline Engineer

Veriipro • Whitpain Township (PA)

On-site
USD 90,000 - 120,000
Senior Java Spark Architect — On-Site Data Pipelines Lead
Senior Java Spark Architect — On-Site Data Pipelines Lead

Tech Mirrors • Cleburne (TX)

On-site
USD 140,000 - 190,000
Senior Big Data Engineer: Hadoop & Spark to Snowflake
Senior Big Data Engineer: Hadoop & Spark to Snowflake

Veriipro • New York (NY)

On-site
USD 140,000 - 190,000
Looking for Java Spark Engineer – Onsite
Looking for Java Spark Engineer – Onsite

Tech Mirrors • Berkeley Heights (NJ)

Hybrid
USD 120,000 - 150,000
Senior Data Engineer: Spark, Java/Scala | Remote GA
Senior Data Engineer: Spark, Java/Scala | Remote GA

EPAM Systems • Georgia

On-site
USD 46,000 - 69,000
Work abroad opportunity
Relocation options
Growth programs
+7
Senior PySpark and Big Data Lead
Senior PySpark and Big Data Lead

Citi • Jersey City (NJ)

On-site
USD 128,000 - 213,000
VP, Big Data & Platform Engineering
VP, Big Data & Platform Engineering

2755 Barclays Services Corpor • Hanover Township (NJ)

On-site
USD 170,000 - 230,000
Incentive award
Competitive benefits