가입 후 초대 링크를 공유하면 동영상 재생 및 초대 보상을 받을 수 있습니다.

Perry Dong
@perryadong
CS PhD @StanfordAILab, Part Time @GoogleDeepMind | Reinforcement Learning
가입 August 2015
114 팔로잉 중    1.2K 팬
World models have emerged as one of the biggest directions in physical AI. At the same time, RL fine-tuning is unlocking capabilities in frontier models beyond what pretraining can achieve on its own Can we get the best of both worlds? We propose Q-Learning with World Models (QWM) (1/7)
더 보기