註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Dwarkesh Patel
@dwarkesh_sp
加入 December 2019
1.1K 正在關注    278.5K 粉絲
Was really interesting to hear John, Beren, and Charlie speculate about why Sonnet 5 and Opus 5 feel like worse models than GLM 5.3 (despite the fact that Anthropic can do raw logit distillation from Fable, and can also train Sonnet/Opus on the environments from which Fable was trained). Led to some interesting thoughts about value of distillation, what it takes to do distillation effectively, and what kinds of model behaviors are hard to extract from distillation.
顯示更多
0
38
1.9K
122
轉發到社區