I always believe LLMs will do the most of "intelligent" part of "embodied intelligence".
We don't need multiple brains.
Robotic models are just physical tools, i.e. infra for LLMs.
Opus 5 can do this with zero demonstrations despite not being trained for robots. On the flip side Opus took 7 minutes while GEN-1.5 took 7 seconds. 🧵