注册并分享邀请链接,可获得视频播放与邀请奖励。

Noam Brown
@polynoamial
Researching reasoning @OpenAI | Co-created Libratus/Pluribus superhuman poker AIs, CICERO Diplomacy AI, and OpenAI o-series 🍓 reasoning models
加入 January 2017
958 正在关注    175.3K 粉丝
@OpenAI o1 is trained with RL to “think” before responding via a private chain of thought. The longer it thinks, the better it does on reasoning tasks. This opens up a new dimension for scaling. We’re no longer bottlenecked by pretraining. We can now scale inference compute too.
显示更多
0
40
2K
213
转发到社区