๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Google DeepMind
@GoogleDeepMind
The engine room of @Google. Building AI safely and responsibly to solve the worldโ€™s most complex problems. Join us:
๊ฐ€์ž… January 2016
276 ํŒ”๋กœ์ž‰ ์ค‘    1.5M ํŒฌ
A modelโ€™s chain of thought acts like a scratch pad, offering a window into its reasoning. ๐Ÿ“ On the latest episode of our podcast, host @fryrsquared sits down with @NeelNanda5 to explore interpretability โ€“ the science of reverse engineering how neural networks learn and think. Timecodes: 00:00 Introduction 02:41 Motivation for interpretability research 04:01 Mechanistic interpretability 08:14 Chain of thought monitoring 18:14 Interpretability techniques 35:00 Auditing models for safety 48:53 What comes next for interpretability
๋” ๋ณด๊ธฐ