“Dream-RSI: Recursive Self-Improvement through Evolving Worlds”
AI agents can search for better solutions, but they’re usually stuck using the same search strategy over and over.
Dream-RSI lets the agent learn how to search better by turning its past exploration into a simulator, where it can cheaply replay different strategies before spending compute in the real world.
So the agent improves not just its solutions, but also the process it uses to discover them.