Register and share your invite link to earn from video plays and referrals.

Shashwat Goel
@ShashwatGoel7
Training AI for Decision Making Past work: Training AI Co-scientists, ΔBelief-RL, Measuring Long Horizon Execution
Joined June 2020
2.3K Following    4.1K Followers
I have to say :) the events were discrete (but yes, noisy!) I don't think the inference from FutureSim should be LLMs suck. The environment is long horizon and relatively open-ended, GPT 5.5 still does surprisingly well. And we know RL makes them better
Show more
@ziv_ravid Continuous, high-dimensional, noisy data. LLMs totally suck at those.