註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Shashwat Goel
@ShashwatGoel7
Training AI for Decision Making Past work: Training AI Co-scientists, ΔBelief-RL, Measuring Long Horizon Execution
加入 June 2020
2.3K 正在關注    4.1K 粉絲
I have to say :) the events were discrete (but yes, noisy!) I don't think the inference from FutureSim should be LLMs suck. The environment is long horizon and relatively open-ended, GPT 5.5 still does surprisingly well. And we know RL makes them better
顯示更多
@ziv_ravid Continuous, high-dimensional, noisy data. LLMs totally suck at those.