登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Jack Morris
@jxmnop
research @engramlab // language models, information theory, science of AI
参加 October 2016
1.1K フォロー中    53.7K ファン
Question for LLM enthusiasts these days, one typically kicks off post training by distilling some reasoning trajectories from some other model. it's kind of like a sourdough starter so how was the *first* model post-trained? did humans write the original set of formatting traces from scratch? (this is why Inkling used a small number of Kimi traces btw)
もっと見る