注册并分享邀请链接,可获得视频播放与邀请奖励。

Jack Morris
@jxmnop
research @engramlab // language models, information theory, science of AI
加入 October 2016
1.1K 正在关注    53.7K 粉丝
Question for LLM enthusiasts these days, one typically kicks off post training by distilling some reasoning trajectories from some other model. it's kind of like a sourdough starter so how was the *first* model post-trained? did humans write the original set of formatting traces from scratch? (this is why Inkling used a small number of Kimi traces btw)
显示更多
0
54
341
8
转发到社区