註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Jack Morris
@jxmnop
research @engramlab // language models, information theory, science of AI
加入 October 2016
1.1K 正在關注    53.7K 粉絲
i'm quite excited about this! • it's cool in general that models can generate their own training data and learn from it. wasn't fully clear to me even a year ago that this would work reliably. • the thing that we're trying to build seems really important and no one has built it before • some traces from our model genuinely surprise me (e.g. the attached example, where it perfectly simulates the output of a complicated bash query, despite not being trained to do this) my main feeling overall is that there is so much to do. we see signs things are starting to work: models use their memories to produce better answers, knowing things saves lots of tokens, and all of this emerges with more compute but solving this memory calibration problem, on top of learning how to generate data in ways that scale nicely with compute, is going to take some time this is just a first step 🫡
顯示更多
0
16
481
32
轉發到社區