Backend Engineer, Distributed Systems

Kloudfuse

Cupertino, Northern (CA, KY)

Hybrid

USD 140,000 - 210,000

Full time

14 days+
Application generator

A complete application in a minute — tailored resume and cover letter, ready to send.

Get past ATS filters

Job summary

Kloudfuse seeks an experienced systems engineer to own the storage and query engine of our observability data lake in Cupertino. You will shape ingestion pipelines, segment layouts, indexing, and tiering across hot and cold storage to keep 300 TB/day responsive.

You’ll push the distributed planning and execution, implement encoding schemes, and design rollups that scale millions of series while maintaining correctness and observability. On-call ownership is part of the design.

Qualifications

  • Deep hands-on experience with distributed data systems such as DBMS, stream processors, or time-series stores.
  • Fluency in fundamentals: consensus, replication, partitioning, consistency models and failure semantics.
  • Performance instincts grounded in measurement and profiling to minimize bytes and cache misses.
  • Comfort with Go or Java at production depth and willingness to work in either.
  • On-call ownership of systems is expected as part of the design feedback.

Responsibilities

  • Own the storage layer of the observability data lake: ingestion pipelines, segment layout, indexing and tiering across hot and cold storage.
  • Push the query engine: distributed planning and execution, predicate pushdown, vectorised scans and cost modeling.
  • Design encodings, sketches and rollups to keep millions of series affordable at scale.
  • Build multi-tenant isolation with quotas, admission control and backpressure that fail predictably.
  • Make correctness observable: consistency guarantees, replay and repair paths, and tests before customer exposure.

Skills

Distributed systems
Go
Java
Production on-call
Observability

Tools

Apache Pinot
Druid
ClickHouse
Kafka
Flink

Job description

Own the storage and query engine underneath the data lake — the part that decides whether 300 TB a day is queryable in seconds or not at all.

What you’ll work on
  • Own the storage layer of the observability data lake: ingestion pipelines, segment layout, indexing, compaction and tiering across hot and cold storage.
  • Push the query engine: distributed planning and execution, predicate pushdown, vectorised scans, and the cost model that decides between them.
  • Hold the line on cardinality. Design the encodings, sketches and rollups that keep millions of series affordable rather than merely possible.
  • Build multi-tenant isolation that holds under load — quotas, admission control and backpressure that fail predictably instead of loudly.
  • Make correctness observable: consistency guarantees, replay and repair paths, and the tests that prove them before a customer does.
What we look for
  • Deep, hands‑on experience with distributed data systems — a database, stream processor, search engine, time‑series store or query engine you helped build and run.
  • Fluency in the fundamentals that decide these systems: consensus, replication, partitioning, consistency models and failure semantics.
  • A performance instinct grounded in measurement — you profile, you read the flamegraph, and you know where the bytes and the cache misses went.
  • Comfort in Go or Java at production depth, and willingness to work in either.
  • Engineers who operate what they build. On‑call for your own system is part of the design feedback, not a chore bolted on afterwards.
Bonus
  • Internals of Apache Pinot, Druid, ClickHouse, Kafka, Flink, Lucene or Parquet — as a contributor or as someone who has debugged them in anger.
  • Columnar and time‑series storage internals: encodings, compression, zone maps, bloom and inverted indexes, cache behaviour.
  • Query language and planner work — parsing, rewriting or optimising PromQL, SQL or something like them.
  • Observability, APM or monitoring domain experience.
How we work
  • Small team, short feedback loops, real ownership from week one.
  • You’ll talk to customers, engineers debugging real incidents, and that shapes what you build.
  • We ship, then iterate; bias toward hands‑on building over process.
  • Modern tooling is encouraged, including AI‑assisted development.
Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Backend Engineer
Backend Engineer

Kloudfuse • Cupertino (CA), Northern (KY)

Hybrid
USD 150,000 - 190,000
Backend/Infra Engineer
Backend/Infra Engineer

Judgment Labs • San Francisco (CA)

On-site
USD 140,000 - 180,000
Staff Software Engineer (Data/Infrastructure)
Staff Software Engineer (Data/Infrastructure)

UMATR • New York (NY)

On-site
USD 225,000 - 275,000
Competitive salary up to $250k plus sizeable equity
Opportunity to influence core architecture
Collaborative, low-bureaucracy culture
Backend/Infra Engineer
Backend/Infra Engineer

Judgment Labs Inc. • San Francisco (CA)

On-site
USD 120,000 - 160,000
Staff Software Engineer (Data/Infrastructure)
Staff Software Engineer (Data/Infrastructure)

UMATR • New York (NY)

On-site
USD 225,000 - 275,000
Competitive salary up to $250k
Sizeable equity
High ownership in architecture
+2
Backend Engineer
Backend Engineer

AnoSys Technologies, Inc • Los Angeles (CA)

On-site
USD 100,000 - 140,000
Middle Software Engineer — Query Engine / Data Platform
Middle Software Engineer — Query Engine / Data Platform

OpportuLand • Northern (KY)

Hybrid
USD 85,000 - 137,000
Senior Full-Stack Engineer: Infrastructure & Distributed Systems
Senior Full-Stack Engineer: Infrastructure & Distributed Systems

3M HEALTHCARE • San Francisco (CA)

On-site
USD 140,000 - 180,000
Senior Software Engineer
Senior Software Engineer

Harnham • Irvine (CA)

On-site
USD 150,000 - 210,000
Competitive base
Bonus
Full benefits
Member of Technical Staff — Data Infrastructure
Member of Technical Staff — Data Infrastructure

Causal Labs • San Francisco (CA)

On-site
USD 150,000 - 210,000