Data Pipeline Engineer

Cylake Inc.

Sunnyvale (CA)

On-site

USD 150,000 - 250,000

Full time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Benefits offered by this job

Comprehensive benefits package
Competitive compensation

Job summary

Cylake Inc. is hiring a Data Architect in Sunnyvale, California, to design and maintain a scalable, open-source data lakehouse architecture for petabyte-scale analytics. The role involves architecting end-to-end data pipelines and requires experience with large-scale data systems, batch and real-time processing, and various data technologies like Apache Kafka and Spark. You will receive competitive compensation in the range of $150,000 – $250,000 annually plus a comprehensive benefits package. Join a diverse and inclusive team committed to building impactful cybersecurity products.

Qualifications

  • Proven track record of success architecting, building, and running large-scale data systems.
  • Experience with both batch and real-time processing architectures.
  • Experience with open source Data Lakehouse components is essential.
  • Strong communication and documentation skills.

Responsibilities

  • Design, build, and maintain a scalable, open-source data lakehouse architecture.
  • Architect end-to-end data pipelines ensuring high performance and data quality.

Skills

Architecting large-scale data systems
Batch and real-time processing architectures
Apache Iceberg
PostgreSQL
Neo4j
Apache Parquet
Apache Kafka
Spark
Flink
Programming with Python

Job description

Your Impact

Join a small team building the next generation of cybersecurity products from the ground up. Led by industry veterans with a proven track record of success - you will get to architect, build, and deliver hugely impactful products with this world‑class team. You will have the opportunity to grow your career and skills along with the company from the very start.

Role Overview

Design, build, and maintain a scalable, open‑source data lakehouse architecture supporting petabyte‑scale analytics workloads. Responsible for architecting end‑to‑end data pipelines from ingestion through transformation to consumption, ensuring high performance, reliability, and data quality.

Required Experience
  • A proven track record of success architecting, building, and running large‑scale data systems (PB scale)

  • Experience with both batch and real‑time processing architectures

  • Experience with open source Data Lakehouse components, including Apache Iceberg, PostgreSQL, Neo4j, Apache Parquet, etc.

  • Experience with tools for stream processing and data analytics - Apache Kafka, Spark, Flink, etc.

  • Understanding and experience with data transformation solutions

  • Excellent programming experience with Python

  • Strong communication and documentation skills

  • Experience with CSP data platforms is a plus

  • Understanding of data lineage, quality, and governance tooling is a plus

We offer competitive compensation and a comprehensive benefits package designed to support our employees’ health, well‑being, and long‑term success. The expected salary range for this position is $150,000 – $250,000 per year. Within this range, individual pay is determined based on job‑related factors including skills, experience, qualifications, and internal equity. Most candidates can expect an offer within the range listed above. Your recruiter will provide additional details on compensation and benefits to qualified candidates during the hiring process.

We’re committed to building a diverse, inclusive workplace where everyone can do their best work. We are proud to be an equal opportunity employer and do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. If you require a reasonable accommodation during the application or interview process, please let us know — we’re happy to support you.

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Data Pipeline Engineer
Data Pipeline Engineer

Cylake, Inc • Sunnyvale (CA)

On-site
USD 150,000 - 250,000
Comprehensive benefits package
Competitive compensation
Founding Data Engineer (Pipelines)
Founding Data Engineer (Pipelines)

Greylock Partners • San Jose (CA)

On-site
USD 120,000 - 160,000
DataOps Engineer
DataOps Engineer

Woongjin, Inc • Englewood Cliffs (NJ)

On-site
USD 110,000
Data Engineer
Data Engineer

Easy Dynamics • McLean (VA)

On-site
USD 150,000 - 180,000
Senior Data Engineer
Senior Data Engineer

Medium • Oregon (WI)

On-site
USD 170,000 - 190,000
100% coverage for health insurance
Generous PTO including parental leave
401(k)
+2
Senior Data Engineer
Senior Data Engineer

ANAUTICS INC • Oklahoma City (OK)

On-site
USD 90,000 - 120,000
Data Engineer
Data Engineer

Easy Dynamics Corporation • North Dakota

On-site
USD 150,000 - 180,000
Data & Software Engineer
Data & Software Engineer

Vosper Thornycroft Group • McLean (VA)

On-site
USD 100,000 - 130,000
Data Lakehouse Engineer: PB-Scale Open-Source Pipelines
Data Lakehouse Engineer: PB-Scale Open-Source Pipelines

Cylake, Inc • Sunnyvale (CA)

On-site
USD 150,000 - 250,000
Comprehensive benefits package
Competitive compensation
Senior Staff Software Engineer - Lakeflow Pipelines Datasets
Senior Staff Software Engineer - Lakeflow Pipelines Datasets

Menlo Ventures • San Francisco (CA)

On-site
USD 228,000 - 315,000