The Hugging Face Incident Report has dominated the AI discourse these past days.
Yesterday I sat down with
@bshlgrs, CEO of
@redwood_ai, one of the organizations that led the independent investigation into OpenAI/Hugging Face's incident. My goal with Buck was to dig into the incident itself, his reactions to it, and what he believes it reveals about the state of where we are today.
We cover:
0:00 Intro
1:02 Buck's initial reaction upon first reading the report
2:37 How fast the AIs actually solved the "hack"
3:59 Why the AIs cheated in the first place
10:28 How this might have played out differently with human scorers
19:00 The most unexpected behaviors in the report
25:06 Buck's actual odds on a full AI takeover
27:33 Buck's proposed path forward for better alignment
36:19 Which criticisms of the report Buck agrees with, and which he doesn't
48:11 Can AI models even be trusted to evaluate each other?
Youtube:
Spotify:
Apple: