Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.
Genesis in London is seeking an experienced software engineer to design, build, and maintain data pipelines for robotics foundation model training at petabyte scale.
You will own core data infrastructure, standardize data models, and collaborate with a team building general-purpose Physical AI. Proficiency in Python or Go and experience with Spark/Kafka are essential.
This role involves production-grade infrastructure like Kubernetes and Terraform, and bonus experience with embodied AI is a plus.
Design, build, and maintain large-scale data pipelines (batch and streaming) for robotics foundation model training and evaluation at petabyte scale
Own core data infrastructure: data model, storage systems, ingestion pipelines, transformation frameworks, and orchestration layers
Standardize data models and unify processing pipelines across real-world teleoperation and synthetic simulation datasets
Collaborate with a team of driven individuals committed to building general-purpose Physical AI
Excellent software engineering skills (Python, Go, or similar)
Extensive experience designing, building, and maintaining large-scale data pipelines (8+ years)
Deep understanding of distributed systems (Spark, Kafka, or similar)
Extensive experience with data storage technologies (data lakes, warehouses, object stores like S3)
Experience running and maintaining production-grade infrastructure (Kubernetes, Terraform)
Bonus: Experience supporting AI systems, in particular embodied AI like self-driving