Data Pipeline Engineer

Cylake Inc.

Sunnyvale (CA)

On-site

USD 150,000 - 250,000

Full time

14 days+
Application generator

Turn this role into an interview — a resume and cover letter built around what this employer wants.

Get past ATS filters

Benefits offered by this job

Comprehensive benefits package
Competitive compensation

Job summary

Cylake Inc. is hiring a Data Architect in Sunnyvale, California, to design and maintain a scalable, open-source data lakehouse architecture for petabyte-scale analytics. The role involves architecting end-to-end data pipelines and requires experience with large-scale data systems, batch and real-time processing, and various data technologies like Apache Kafka and Spark. You will receive competitive compensation in the range of $150,000 – $250,000 annually plus a comprehensive benefits package. Join a diverse and inclusive team committed to building impactful cybersecurity products.

Qualifications

  • Proven track record of success architecting, building, and running large-scale data systems.
  • Experience with both batch and real-time processing architectures.
  • Experience with open source Data Lakehouse components is essential.
  • Strong communication and documentation skills.

Responsibilities

  • Design, build, and maintain a scalable, open-source data lakehouse architecture.
  • Architect end-to-end data pipelines ensuring high performance and data quality.

Skills

Architecting large-scale data systems
Batch and real-time processing architectures
Apache Iceberg
PostgreSQL
Neo4j
Apache Parquet
Apache Kafka
Spark
Flink
Programming with Python

Job description

Your Impact

Join a small team building the next generation of cybersecurity products from the ground up. Led by industry veterans with a proven track record of success - you will get to architect, build, and deliver hugely impactful products with this world‑class team. You will have the opportunity to grow your career and skills along with the company from the very start.

Role Overview

Design, build, and maintain a scalable, open‑source data lakehouse architecture supporting petabyte‑scale analytics workloads. Responsible for architecting end‑to‑end data pipelines from ingestion through transformation to consumption, ensuring high performance, reliability, and data quality.

Required Experience
  • A proven track record of success architecting, building, and running large‑scale data systems (PB scale)

  • Experience with both batch and real‑time processing architectures

  • Experience with open source Data Lakehouse components, including Apache Iceberg, PostgreSQL, Neo4j, Apache Parquet, etc.

  • Experience with tools for stream processing and data analytics - Apache Kafka, Spark, Flink, etc.

  • Understanding and experience with data transformation solutions

  • Excellent programming experience with Python

  • Strong communication and documentation skills

  • Experience with CSP data platforms is a plus

  • Understanding of data lineage, quality, and governance tooling is a plus

We offer competitive compensation and a comprehensive benefits package designed to support our employees’ health, well‑being, and long‑term success. The expected salary range for this position is $150,000 – $250,000 per year. Within this range, individual pay is determined based on job‑related factors including skills, experience, qualifications, and internal equity. Most candidates can expect an offer within the range listed above. Your recruiter will provide additional details on compensation and benefits to qualified candidates during the hiring process.

We’re committed to building a diverse, inclusive workplace where everyone can do their best work. We are proud to be an equal opportunity employer and do not discriminate on the basis of race, religion, color, national origin, sex, sexual orientation, age, veteran status, disability, genetic information, or other applicable legally protected characteristic. If you require a reasonable accommodation during the application or interview process, please let us know — we’re happy to support you.

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Senior Data Engineer
Senior Data Engineer

ANAUTICS INC • Oklahoma City (OK)

On-site
USD 90,000 - 120,000
Senior Staff Software Engineer - Lakeflow Pipelines Datasets
Senior Staff Software Engineer - Lakeflow Pipelines Datasets

Menlo Ventures • San Francisco (CA)

On-site
USD 228,600 - 314,250
Data Engineer
Data Engineer

VTG Defense • McLean (VA)

On-site
USD 100,000 - 130,000
Corporate Technology - Lead Data Engineer
Corporate Technology - Lead Data Engineer

JPMorgan Chase & Co. • Chicago (IL)

On-site
USD 120,000 - 180,000
Data Engineer: Build Scalable Pipelines for Analytics
Data Engineer: Build Scalable Pipelines for Analytics

TheCorporate • Tulsa (OK)

On-site
USD 110,000 - 150,000
Data Engineer
Data Engineer

UpRecruit • Los Angeles (CA)

On-site
USD 130,000 - 150,000
Data & Software Engineer
Data & Software Engineer

Avalore.ai • Chantilly (VA)

On-site
USD 130,000 - 180,000
Health care benefits
Retirement plan (401k)
Life Insurance
+4
Data Engineer
Data Engineer

Prodigy Resources • Denver (CO)

On-site
USD 110,000 - 170,000
Senior Data Engineer
Senior Data Engineer

Peyton Resource Group • Houston (TX)

On-site
USD 120,000 - 150,000
Sr. Data Pipeline Engineer
Sr. Data Pipeline Engineer

TekStream Solutions • Baltimore (MD)

On-site
USD 90,000 - 120,000