注册并分享邀请链接,可获得视频播放与邀请奖励。

Jack Morris
@jxmnop
research @engramlab // language models, information theory, science of AI
加入 October 2016
1.1K 正在关注    53.7K 粉丝
i'm quite excited about this! • it's cool in general that models can generate their own training data and learn from it. wasn't fully clear to me even a year ago that this would work reliably. • the thing that we're trying to build seems really important and no one has built it before • some traces from our model genuinely surprise me (e.g. the attached example, where it perfectly simulates the output of a complicated bash query, despite not being trained to do this) my main feeling overall is that there is so much to do. we see signs things are starting to work: models use their memories to produce better answers, knowing things saves lots of tokens, and all of this emerges with more compute but solving this memory calibration problem, on top of learning how to generate data in ways that scale nicely with compute, is going to take some time this is just a first step 🫡
显示更多
0
16
481
32
转发到社区