Reproducible by default.
We release the benchmarks, datasets, and software behind our papers, so that the results can be checked and built on by others. Each release below is tied to a publication where one exists.
The reference implementation of our perceive–plan–act stack: a shared representation layer with perception, planning, and control back ends that read and write the same belief state.
A benchmark for evaluating embodied cognition systems end to end — measuring how errors in one stage propagate through the full loop rather than scoring each stage in isolation.
A dataset of manipulation trajectories recorded on hardware across insertion, assembly, and reorientation tasks, with per-frame contact and proprioceptive annotations for training world models.
A domain-randomisation and adversarial-distillation toolkit for training policies in simulation and transferring them to hardware with minimal task-specific tuning.