注册并分享邀请链接,可获得视频播放与邀请奖励。

Baseten
@baseten
Inference is everything.
加入 March 2021
81 正在关注    19.8K 粉丝
RL teaches models to work longer, but reasoning is dependent on domain-specific post-training. Baseten's Head of Model Training @oneill_c sat down with @dwarkesh_sp to explain horizon generalization and what's next at the frontier. Full episode here:
显示更多