Here’s what I’m watching tonight: 4 hours of Figure 03 robots doing home chores.
Long, uncut scenes of housework in 30 different homes.
Figure introduced Helix 2.5, its most advanced humanoid AI yet, built to answer one question: can a humanoid walk into a home it has never seen and get to work on its own?
The setup
One foundation model pretrained on Index, Figure’s global dataset of human behavior, then fine-tuned into three whole-body behaviors: tidying living rooms, folding towels, and making beds. Tested across 30 Bay Area homes with zero data collected in any of them, and with objects it had never seen.
Results
• 56% zero-shot success vs. 9% trained from scratch. Index pretraining was the only variable.
• No partial credit in success rate. Every toy in the basket, every towel folded, the whole bed made.
• Half the task data of a comparable Helix 02 behavior, and 30× wider generalization.
• A human-to-robot transfer scaling law: doubling Index data improved action prediction predictably enough to forecast the largest run’s loss to four decimals.
• Whole-body self-correction. The robot steps back, repositions, and walks around the bed to fix a fold.
Index now generates ~35 minutes of new human experience every second, and Figure has committed $3.5B of compute to Helix.