註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Google DeepMind
@GoogleDeepMind
The engine room of @Google. Building AI safely and responsibly to solve the world’s most complex problems. Join us:
加入 January 2016
276 正在關注    1.5M 粉絲
A model’s chain of thought acts like a scratch pad, offering a window into its reasoning. 📝 On the latest episode of our podcast, host @fryrsquared sits down with @NeelNanda5 to explore interpretability – the science of reverse engineering how neural networks learn and think. Timecodes: 00:00 Introduction 02:41 Motivation for interpretability research 04:01 Mechanistic interpretability 08:14 Chain of thought monitoring 18:14 Interpretability techniques 35:00 Auditing models for safety 48:53 What comes next for interpretability
顯示更多
0
45
648
91
轉發到社區