注册并分享邀请链接,可获得视频播放与邀请奖励。

Shashwat Goel
@ShashwatGoel7
Training AI for Decision Making Past work: Training AI Co-scientists, ΔBelief-RL, Measuring Long Horizon Execution
加入 June 2020
2.3K 正在关注    4.1K 粉丝
I have to say :) the events were discrete (but yes, noisy!) I don't think the inference from FutureSim should be LLMs suck. The environment is long horizon and relatively open-ended, GPT 5.5 still does surprisingly well. And we know RL makes them better
显示更多
@ziv_ravid Continuous, high-dimensional, noisy data. LLMs totally suck at those.