The dataset and paper based on the previous phase of community contributions are now live:
Paper Link:
Project Page:
Dataset Link:
Github Codebase:
Using Pi0.5 + AXIS-100%, we achieved 88.8 overall success on LIBERO-Plus.
For comparison:
Vanilla Pi0.5 achieved 83.9
A RoboCasa-matched simulation baseline achieved 57.5
This clearly demonstrates that diverse, in-the-wild data provides substantial gains in model performance.
And V2 is already in progress — significantly larger in scale, covering more embodiments and a wider range of atomic capabilities.