Java Spark Engineer

Saransh Inc

Berkley (AL)

On-site

USD 150,000 - 210,000

Full time

8 days ago
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Job summary

Saransh Inc. seeks a Java Spark Engineer to architect and implement scalable data pipelines using Apache Spark in production. You will lead batch and streaming ETL/ELT designs, optimize performance, and set coding standards across a growing team.

Located on-site in Berkley Heights, NJ, the role requires 10+ years of experience, deep distributed systems knowledge, and strong SQL with columnar formats. Collaboration with product and analytics teams is essential.

Qualifications

  • Bachelor's or Master's degree in Computer Science, Engineering, or related field.
  • 7+ years of professional Java development experience.
  • 5+ years hands-on experience with Apache Spark in production environments.
  • Expert-level understanding of distributed systems: fault tolerance, data locality, shuffle mechanics, resource management.
  • Proven track record designing systems processing terabyte+ scale data.
  • Strong SQL skills and deep familiarity with columnar storage formats (Parquet, ORC, Avro, Delta Lake).
  • Experience with cluster managers (YARN, Kubernetes) and cloud-managed Spark.
  • Proficiency with Kafka.
  • Strong grasp of CI/CD, containerization, and infrastructure-as-code practices.

Responsibilities

  • Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java)
  • Lead design of batch and streaming ETL/ELT systems handling large data volumes
  • Deep-dive performance tuning: partitioning strategy, memory management, shuffle/skew optimization, job cost reduction
  • Set coding standards and lead code/design reviews across the team
  • Drive technical decisions on data architecture, storage formats, and pipeline orchestration
  • Mentor mid-level and junior engineers; act as a technical escalation point
  • Partner with product, analytics, and platform teams to translate requirements into scalable systems
  • Own production reliability - on-call ownership, incident response, root-cause analysis for pipeline failures
  • Evaluate and introduce new tools/frameworks where they improve the system
  • Contribute to capacity planning and cost optimization for cluster infrastructure

Skills

Java development
Apache Spark
Distributed systems
SQL
Kafka
CI/CD
Containerization
Orchestration

Education

BSc/ MSc CS or related

Tools

YARN
Kubernetes
Parquet/ORC/Avro/Delta
Delta Lake
Kafka

Job description

Role:Java Spark Engineer

Location:Berkley Heights, NJ (Onsite)

Experience: 10+ years

Responsibilities:

Architect and build scalable, fault-tolerant data pipelines using Apache Spark (Java)

Lead design of batch and streaming ETL/ELT systems handling large data volumes

Deep-dive performance tuning: partitioning strategy, memory management, shuffle/skew optimization, job cost reduction

Set coding standards and lead code/design reviews across the team

Drive technical decisions on data architecture, storage formats, and pipeline orchestration

Mentor mid-level and junior engineers; act as a technical escalation point

Partner with product, analytics, and platform teams to translate requirements into scalable systems

Own production reliability - on-call ownership, incident response, root-cause analysis for pipeline failures

Evaluate and introduce new tools/frameworks where they improve the system

Contribute to capacity planning and cost optimization for cluster infrastructure

Required Qualifications

Bachelor's or Master's degree in Computer Science, Engineering, or related field

7+ years of professional Java development experience

5+ years hands-on experience with Apache Spark in production environments

Expert-level understanding of distributed systems: fault tolerance, data locality, shuffle mechanics, resource management

Proven track record designing systems processing terabyte+ scale data

Strong SQL skills and deep familiarity with columnar storage formats (Parquet, ORC, Avro, Delta Lake/Iceberg)

Experience with cluster managers (YARN, Kubernetes) and cloud-managed Spark

Proficiency with Kafka

Strong grasp of CI/CD, containerization, and infrastructure-as-code practices

Preferred Qualifications

Experience with Flink or other stream-processing frameworks

Familiarity with data governance, lineage, and quality frameworks

Experience with workflow orchestration at scale

Background in system design for multi-tenant or multi-region data platforms

Prior experience leading a team or acting as a technical lead

Soft Skills / Leadership

Excellent communication - able to explain technical tradeoffs to non-technical stakeholders

Strong mentorship and coaching ability

Comfortable driving ambiguous, cross-team technical initiatives

Behavioral Skills:

  • Good Communication skills
  • 5 days Work from Office at Berkley Heights, NJ
  • Team Player
  • Ability to work in a changing environment
  • Strong problem solving and analytical skills
  • Ability to work independently or within a team
  • Manage day-to-day challenges and communicate developmental risks with the technical team
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Java Spark Engineer
Java Spark Engineer

Delta System & Software, Inc. • Berkeley Heights (NJ)

On-site
USD 150,000 - 210,000
Senior Software Engineer - Core Java & Apache Spark
Senior Software Engineer - Core Java & Apache Spark

Citi • New York (NY)

On-site
USD 120,000 - 160,000
Senior Data Engineer – Enterprise Data Frameworks
Senior Data Engineer – Enterprise Data Frameworks

Citizens • Rhode Island

On-site
USD 120,000 - 160,000
Senior Java Spark Engineer Onsite: Data Pipeline Architect
Senior Java Spark Engineer Onsite: Data Pipeline Architect

Saransh Inc • Berkley (AL)

On-site
USD 150,000 - 210,000
Apache Spark Developer
Apache Spark Developer

Bright Vision Technologies • Austin (TX)

Remote
USD 125,000 - 185,000
Apache Spark Java Developer
Apache Spark Java Developer

Elite Workforce Inc. • Brooklyn Park (MN)

On-site
USD 80,000 - 110,000
Software Engineering - Core Java - Intermediate
Software Engineering - Core Java - Intermediate

Ampcus Inc • Chicago (IL)

On-site
USD 110,000 - 150,000
Data Engineer ETL
Data Engineer ETL

Compunnel, Inc. • Durham (NC)

On-site
USD 100,000 - 130,000
Senior PySpark & Databricks Data Engineer
Senior PySpark & Databricks Data Engineer

InfoSmart Technologies, Inc • Atlanta (GA)

Hybrid
USD 120,000 - 150,000
Senior Java Spark Architect & Data Pipeline Tech Lead
Senior Java Spark Architect & Data Pipeline Tech Lead

Mphasis • New York (NY)

On-site
USD 63,000 - 135,000