注册并分享邀请链接,可获得视频播放与邀请奖励。

Google DeepMind
@GoogleDeepMind
The engine room of @Google. Building AI safely and responsibly to solve the world’s most complex problems. Join us:
加入 January 2016
275 正在关注    1.5M 粉丝
A model’s chain of thought acts like a scratch pad, offering a window into its reasoning. 📝 On the latest episode of our podcast, host @fryrsquared sits down with @NeelNanda5 to explore interpretability – the science of reverse engineering how neural networks learn and think. Timecodes: 00:00 Introduction 02:41 Motivation for interpretability research 04:01 Mechanistic interpretability 08:14 Chain of thought monitoring 18:14 Interpretability techniques 35:00 Auditing models for safety 48:53 What comes next for interpretability
显示更多
0
45
648
91
转发到社区