註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Baseten
@baseten
Inference is everything.
加入 March 2021
81 正在關注    19.8K 粉絲
RL teaches models to work longer, but reasoning is dependent on domain-specific post-training. Baseten's Head of Model Training @oneill_c sat down with @dwarkesh_sp to explain horizon generalization and what's next at the frontier. Full episode here:
顯示更多