Google DeepMind just published Dream-RSI....
it recursive self improvement through evolving worlds
not the model rewriting itself into God but the explorer learning how to search
coding agents waste a fortune trying kernels, algorithms, math... fixed search goes stale and tuning search online means paying the agent again
Dream-RSI treats the discovery tree the agent already built as a world replay that world... test thousands of new search policies with zero new runs
deploy the winner.... It finds new stuff that tree joins the pool and it Loops