註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Noam Brown
@polynoamial
Researching reasoning @OpenAI | Co-created Libratus/Pluribus superhuman poker AIs, CICERO Diplomacy AI, and OpenAI o-series 🍓 reasoning models
加入 January 2017
958 正在關注    175.3K 粉絲
@OpenAI o1 is trained with RL to “think” before responding via a private chain of thought. The longer it thinks, the better it does on reasoning tasks. This opens up a new dimension for scaling. We’re no longer bottlenecked by pretraining. We can now scale inference compute too.
顯示更多
0
40
2K
213
轉發到社區