注册并分享邀请链接,可获得视频播放与邀请奖励。

Dwarkesh Patel
@dwarkesh_sp
加入 December 2019
1.1K 正在关注    278.8K 粉丝
Was really interesting to hear John, Beren, and Charlie speculate about why Sonnet 5 and Opus 5 feel like worse models than GLM 5.3 (despite the fact that Anthropic can do raw logit distillation from Fable, and can also train Sonnet/Opus on the environments from which Fable was trained). Led to some interesting thoughts about value of distillation, what it takes to do distillation effectively, and what kinds of model behaviors are hard to extract from distillation.
显示更多
0
38
1.9K
122
转发到社区