登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

alphaXiv
@askalphaxiv
High fidelity research
参加 November 2023
101 フォロー中    56.9K ファン
“Dream-RSI: Recursive Self-Improvement through Evolving Worlds” AI agents can search for better solutions, but they’re usually stuck using the same search strategy over and over. Dream-RSI lets the agent learn how to search better by turning its past exploration into a simulator, where it can cheaply replay different strategies before spending compute in the real world. So the agent improves not just its solutions, but also the process it uses to discover them.
もっと見る