Staff Software Engineer - Apache Spark

Cloudera

United States

On-site

USD 184,000 - 230,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

Cloudera is seeking a Staff Software Engineer to architect and build next-generation distributed systems for the Apache Spark team. You will own features at massive scale, contribute to open-source Spark, and work with a globally distributed team across Data Engineering.

The role emphasizes deep expertise in distributed systems, JVM languages, and collaboration to deliver enterprise-grade data platforms. This position targets senior engineers ready to tackle big data challenges.

Qualifications

  • 5-7+ years of professional software development experience.
  • Proven leadership in technical initiatives and delivering complex product enhancements.
  • Strong proficiency in Java or Scala and other JVM-based languages.
  • Solid experience designing and developing distributed systems.
  • Clear communication for collaboration across distributed teams.

Responsibilities

  • Architect, implement, and deliver next‑generation features for Cloudera's Data Engineering Experience at massive scale.
  • Contribute to Apache Spark and shape the future of distributed data processing.
  • Develop high-performance features using Java, Scala, and Python on modern data platforms.
  • Deepen expertise in core distributed data processing concepts and stack components.
  • Collaborate with a distributed team to drive product vision and delivery.

Skills

Java
Scala
Distributed systems
Leadership
Communication
Autonomy

Job description

Business Area: Engineering

Seniority Level: Mid-Senior level

Job Description

At Cloudera, we empower people to transform complex data into clear and actionable insights. With as much data under management as the hyperscalers, we're the preferred data partner for the top companies in almost every industry. Powered by the relentless innovation of the open source community, Cloudera advances digital transformation for the world's largest enterprises.

The Data Platform Pillar is the bedrock of Cloudera's technology, where we design and build the core components that let our customers store, manage, and process data with unmatched scalability, security, and performance.

Are you ready to architect the future of big data? Cloudera is searching for a visionary Staff Software Engineer with deep expertise in distributed systems to join the Apache Spark Team. You will be at the forefront of innovation, building our next-generation, enterprise-grade system designed to conquer data challenges at a massive scale-running Spark on thousands of nodes and crunching petabytes of data for the world's largest companies. This is your chance to directly influence the open-source community as a key contributor to Apache Spark while collaborating with a high-impact, distributed team that includes multiple Spark committers. If you're passionate about pushing the boundaries of distributed data processing, come build the impossible with us.

As a Staff Engineer you will:
  • Pioneer Scalable Solutions: Architect, implement, and deliver next-generation features for Cloudera's Data Engineering Experience, operating at a massive scale on thousands of production nodes.

  • Drive Open-Source Innovation: Be a core contributor to Apache Spark, directly shaping the future of distributed data processing in the open-source community.

  • Build with Modern Stacks: Develop high-performance features using Scala, Java, and Python on modern data platforms.

  • Deepen Technical Mastery: Gain and apply expert-level knowledge in core distributed data processing concepts, including:

    • SQL Planners and Optimizers

    • Data layout and modern table formats like Apache Parquet and Iceberg

    • Fault tolerance and resilience in large-scale distributed systems

  • Own the Technology Stack: Develop a deep technical understanding of components across the Cloudera Data Engineering Experience, with a focus on Iceberg and Spark, applying this knowledge to your daily tasks.

  • Conquer Large-Scale Challenges: Work hands-on with massive distributed systems, scaling from hundreds to thousands of nodes in live production clusters.

  • Ensure System Integrity: Conduct thorough root cause analysis, debug complex system-level deployment issues, and resolve failures to maintain high system quality.

  • Enhance Engineering Velocity: Improve internal infrastructure and tooling to streamline development, testing, and deployment processes.

  • Collaborate and Influence: Work closely with a high-impact, distributed team and stakeholders to drive product vision and delivery.

We are excited about you if you have:
  • Professional Experience: 5-7+ years of experience in professional software development.

  • Leadership & Delivery: Proven experience leading technical initiatives and delivering complex product enhancements from concept to production.

  • Core Languages: Strong proficiency in Java, Scala, or other JVM-based language.

  • Systems Expertise: Solid experience in the design and development of distributed systems.

  • Engineering Excellence: Passion for clean coding, attention to detail, and a focus on software quality and maintainability.

  • Communication: Strong oral and written communication skills for effective collaboration across a distributed team.

  • Autonomy: Demonstrated ability to research, problem-solve, and operate independently without constant supervision.

  • Growth Mindset: An open-minded approach with a desire to learn new technologies and an unwavering passion for building exceptional products.

You might also have:
  • Spark & Ecosystem Experience with using/developing Apache Spark, Apache Iceberg, or other related technologies.

  • Distributed Systems Mastery: Deep experience with large-scale, distributed systems design and development, including a strong understanding of scaling, performance optimization, and scheduling.

  • SQL Expertise: Experience with SQL Planners and Optimizers

  • Open-Source Contributions: Prior experience as a contributor to open-source projects.

Why this role matters:

You will tackle complex distributed systems challenges, crafting the foundational software for the control and data planes that powers CDP and keeps it running at massive scale. Working at the forefront of hybrid and multi-cloud technology, you will empower data scientists, engineers, and analysts with the tools and infrastructure they need for advanced analytics and modeling.

Collaboration is key, you will work alongside brilliant minds across product, data science, and engineering to drive innovation, standardize best practices, and shape the future of enterprise AI and data platforms. This is your chance to build the future of data and see your work make a global impact.

The expected base salary range for this role in:

  • California & Washington is $184,000 - $230,000 USD

  • Canada is

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Staff Software Engineer - Apache Spark
Staff Software Engineer - Apache Spark

Cloudera • Richmond (VA)

On-site
USD 184,000 - 230,000
Generous PTO Policy
Flexible WFH Policy
Mental & Physical Wellness programs
Staff Software Engineer - Apache Spark
Staff Software Engineer - Apache Spark

Cloudera • California (MO)

Hybrid
USD 184,000 - 230,000
Generous PTO Policy
Flexible WFH Policy
Mental & Physical Wellness programs
Sr. Staff Software Engineer - Apache Iceberg
Sr. Staff Software Engineer - Apache Iceberg

Cloudera • Washington

On-site
USD 184,000 - 230,000
Generous PTO
Unplugged days
Flexible WFH
+6
Staff Software Engineer—Spark at Scale
Staff Software Engineer—Spark at Scale

Cloudera • California (MO)

Hybrid
USD 184,000 - 230,000
Generous PTO Policy
Flexible WFH Policy
Mental & Physical Wellness programs
Staff Spark Engineer: Scale Big Data (Remote)
Staff Spark Engineer: Scale Big Data (Remote)

Cloudera • United States

On-site
USD 184,000 - 230,000
Generous PTO Policy
Flexible WFH Policy
Mental & Physical Wellness programs
+1
Staff Spark Engineer: Scale Massive Data & Open Source
Staff Spark Engineer: Scale Massive Data & Open Source

Cloudera • United States

On-site
USD 184,000 - 230,000
Staff Software Engineer - Spark at Scale (Open Source)
Staff Software Engineer - Spark at Scale (Open Source)

Cloudera • Richmond (VA)

On-site
USD 184,000 - 230,000
Generous PTO Policy
Flexible WFH Policy
Mental & Physical Wellness programs
Sr. Staff Software Engineer - Apache Iceberg
Sr. Staff Software Engineer - Apache Iceberg

Cloudera • Washington

Hybrid
USD 184,000 - 230,000
Generous PTO Policy
Flexible WFH Policy
Mental & Physical Wellness programs
+1
Staff Software Engineer
Staff Software Engineer

Cloudera • Town of Texas (WI)

On-site
USD 150,000 - 230,000
Data Engineer
Data Engineer

Tata Consultancy Services • Irving (TX)

On-site
USD 125,000 - 140,000