注册并分享邀请链接,可获得视频播放与邀请奖励。

Michael Y. Li
@michaelyli_
CS PhD @StanfordAILab @StanfordNLP advised by @noahdgoodman and Emily Fox. Prev: undergrad @princeton
加入 March 2024
497 正在关注    3K 粉丝
You're wasting FLOPs when scaling inference compute: by independently sampling parallel attempts, you burn compute rediscovering the same solutions. Introducing QuasiMoTTo: we scale parallel sampling with correlated samples instead! These samples have higher coverage, are marginally exact draws from the LLM, and can be generated in parallel. Result: same performance with 25-47% fewer samples in test-time scaling + 50% fewer training steps in RL! In our new paper, we explore the design space of correlated samplers. Work with co-authors @probablynotaz9 (co-lead), @gandhikanishk, @noahdgoodman, and Emily Fox!
显示更多
0
14
284
69
转发到社区