註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

alphaXiv
@askalphaxiv
High fidelity research
加入 November 2023
101 正在關注    56.9K 粉絲
“Dream-RSI: Recursive Self-Improvement through Evolving Worlds” AI agents can search for better solutions, but they’re usually stuck using the same search strategy over and over. Dream-RSI lets the agent learn how to search better by turning its past exploration into a simulator, where it can cheaply replay different strategies before spending compute in the real world. So the agent improves not just its solutions, but also the process it uses to discover them.
顯示更多
0
7
171
31
轉發到社區