It's so satisfying to watch the natural, dexterous behavior that our Play2Perfect policy produces.
Play2Perfect first learns a base policy with generic skills like grasping, in-hand re-orientation and goal reaching.
Then it finetunes this policy for a specific assembly task.
🤖 How can we teach dexterous robots to perform precise, contact-rich assembly?
Introducing Play2Perfect: first learn to play with objects, then perfect the policy for tight insertion, multi-part assembly, and screwing.
Sound on! 🔊
🧵👇