註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Ismail Labiad
@is_labiad
加入 June 2025
37 正在關注    39 粉絲
Repeated sampling is the default way to scale LLM reasoning at test time. But token level noise often produces many near duplicate attempts that follow the same high level idea. 🧵 How can we cover more of the solution space without sacrificing throughput?
顯示更多