가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Shashwat Goel
@ShashwatGoel7
Training AI for Decision Making Past work: Training AI Co-scientists, ΔBelief-RL, Measuring Long Horizon Execution
가입 June 2020
2.3K 팔로잉 중    4.1K
I have to say :) the events were discrete (but yes, noisy!) I don't think the inference from FutureSim should be LLMs suck. The environment is long horizon and relatively open-ended, GPT 5.5 still does surprisingly well. And we know RL makes them better
더 보기
@ziv_ravid Continuous, high-dimensional, noisy data. LLMs totally suck at those.