注册并分享邀请链接,可获得视频播放与邀请奖励。

alphaXiv
@askalphaxiv
High fidelity research
加入 November 2023
101 正在关注    56.9K 粉丝
“Dream-RSI: Recursive Self-Improvement through Evolving Worlds” AI agents can search for better solutions, but they’re usually stuck using the same search strategy over and over. Dream-RSI lets the agent learn how to search better by turning its past exploration into a simulator, where it can cheaply replay different strategies before spending compute in the real world. So the agent improves not just its solutions, but also the process it uses to discover them.
显示更多
0
7
171
31
转发到社区