Register and share your invite link to earn from video plays and referrals.

Google DeepMind
@GoogleDeepMind
The engine room of @Google. Building AI safely and responsibly to solve the world’s most complex problems. Join us:
Joined January 2016
276 Following    1.5M Followers
A model’s chain of thought acts like a scratch pad, offering a window into its reasoning. 📝 On the latest episode of our podcast, host @fryrsquared sits down with @NeelNanda5 to explore interpretability – the science of reverse engineering how neural networks learn and think. Timecodes: 00:00 Introduction 02:41 Motivation for interpretability research 04:01 Mechanistic interpretability 08:14 Chain of thought monitoring 18:14 Interpretability techniques 35:00 Auditing models for safety 48:53 What comes next for interpretability
Show more