Get a reply from this employer — a resume and cover letter tailored to exactly what they’re hiring for.
Nebius B.V. is seeking a Software Engineer with strong C++ expertise to join the Nebius Data Platform team.
You will design and implement core platform features, ensure reliability, and work across a multi-tenant system used for business-critical pipelines, ML workloads, and analytics. You’ll collaborate with platform engineers across Go and Python services, contribute to scalable storage and compute components, and own production quality through monitoring, incident response, and thoughtful
We’re looking for a Software Engineer with strong C++ expertise to join the team building and operating Nebius Data Platform — a distributed storage and a processing platform that acts as the company’s “source of truth” and the backbone of many internal (and some external) products. Nebius Data Platform is a single multi-tenant ecosystem based on YTsaurus — instead of running separate HDFS/Kafka/HBase-style systems, we provide storage, compute, and analytics capabilities inside one platform. Built on top of the open-source YTsaurus ecosystem, we run and extend our own Nebius distribution and develop significant in-house functionality (core and platform-level). We can design, implement, and roll out features end-to-end on our clusters without waiting for upstream approvals and contribute upstream when it makes sense. At scale today, this includes ~500 servers, ~20k CPU cores and ~10 PB of compressed data in our largest production cluster, supporting workloads ranging from business-critical pipelines and financial transactions to large-scale ML/LLM training datasets and compute.
You’ll work on a system that includes (and ties together): Distributed Storage (Cypress): transactional semantics, tiered storage, erasure coding, replication, and strong reliability expectations. Compute & ETL: a cluster-wide job scheduler (tens of thousands of cores), MapReduce, YQL for SQL-like data processing, and SPYT (Spark over YTsaurus) for modern data engineering. Interactive analytics (CHYT): ClickHouse® instances spun up directly on compute nodes for fast SQL over data in-place. Dynamic Tables: low-latency NoSQL KV with distributed ACID transactions for OLTP-style workloads and feature stores. Orchestracto: workflow orchestration deeply integrated with the platform (Airflow-like, but platform-native).
We’re looking for engineers who combine strong systems skills with product sense: understanding who uses the platform, why certain capabilities matter, and making pragmatic trade-offs to maximize impact. On our team, engineering work is expected to be connected to real users and outcomes — you’ll regularly align with internal stakeholders, clarify requirements, and help drive prioritization.
5+ years of software engineering experience. Strong C++ skills (you’ll write core code). Working knowledge of Python and/or Go (you don’t have to be expert, but should be comfortable navigating them). Experience developing and/or operating high‑load, distributed services. Production mindset: ability to use SSH, read logs/metrics/traces, and debug distributed systems behavior. Solid CS fundamentals: algorithms, data structures, concurrency basics.
We conduct coding interviews as part of the process.
5+ years of software engineering experience, Strong C++ skills (core code), Working knowledge of Python and/or Go, Experience developing and/or operating high-load, distributed services, Production mindset: SSH, logs/metrics/traces, debugging distributed systems, Solid CS fundamentals: algorithms, data structures, concurrency basics