註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Perry Dong
@perryadong
CS PhD @StanfordAILab, Part Time @GoogleDeepMind | Reinforcement Learning
加入 August 2015
114 正在關注    1.2K 粉絲
World models have emerged as one of the biggest directions in physical AI. At the same time, RL fine-tuning is unlocking capabilities in frontier models beyond what pretraining can achieve on its own Can we get the best of both worlds? We propose Q-Learning with World Models (QWM) (1/7)
顯示更多
0
7
515
56
轉發到社區