
Data Infrastructure at Genesis (London, United Kingdom)
Genesis· London, United Kingdom·
Role details
Work type
Hybrid
Employment
Full-Time
Job description
What You’ll Do
- Design, build, and maintain large-scale data pipelines (batch and streaming) for robotics foundation model training and evaluation at petabyte scale
- Own core data infrastructure: data model, storage systems, ingestion pipelines, transformation frameworks, and orchestration layers
- Standardize data models and unify processing pipelines across real-world teleoperation and synthetic simulation datasets
- Collaborate with a team of driven individuals committed to building general-purpose Physical AI
What You’ll Bring
- Excellent software engineering skills (Python, Go, or similar)
- Extensive experience designing, building, and maintaining large-scale data pipelines (8+ years)
- Deep understanding of distributed systems (Spark, Kafka, or similar)
- Extensive experience with data storage technologies (data lakes, warehouses, object stores like S3)
- Experience running and maintaining production-grade infrastructure (Kubernetes, Terraform)
- Bonus: Experience supporting AI systems, in particular embodied AI like self-driving
Why work at Genesis
- Culture: Values include candor, responsibility, limitless ambition, and a joyful journey. The team is described as multicultural, optimistic, and pragmatic.
- Work model: Hybrid/office-based presence in Paris, San Francisco Bay Area, and London. No explicit remote policy mentioned, but the company emphasizes in-person collaboration.
- Engineering culture: Builders get to work on the hardest problems in robotics and AI – from foundation models to hardware design to data systems. The team includes pioneers of generative simulation, Diffusion Policy, and GPU compilers.
- Notable perks: The chance to shape a category-defining general-purpose robot; close collaboration with world-class investors and advisors; fast-growing startup with significant resources ($105M seed).