注册并分享邀请链接,可获得视频播放与邀请奖励。

Ismail Labiad
@is_labiad
加入 June 2025
37 正在关注    39 粉丝
Repeated sampling is the default way to scale LLM reasoning at test time. But token level noise often produces many near duplicate attempts that follow the same high level idea. 🧵 How can we cover more of the solution space without sacrificing throughput?
显示更多