註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Michael Y. Li
@michaelyli_
CS PhD @StanfordAILab @StanfordNLP advised by @noahdgoodman and Emily Fox. Prev: undergrad @princeton
加入 March 2024
497 正在關注    3K 粉絲
You're wasting FLOPs when scaling inference compute: by independently sampling parallel attempts, you burn compute rediscovering the same solutions. Introducing QuasiMoTTo: we scale parallel sampling with correlated samples instead! These samples have higher coverage, are marginally exact draws from the LLM, and can be generated in parallel. Result: same performance with 25-47% fewer samples in test-time scaling + 50% fewer training steps in RL! In our new paper, we explore the design space of correlated samplers. Work with co-authors @probablynotaz9 (co-lead), @gandhikanishk, @noahdgoodman, and Emily Fox!
顯示更多
0
14
284
69
轉發到社區