註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

MCNAIR
@mcnairai
We mitigate catastrophic loss-of-control risks from advanced AI through low-effort, high-impact research. Posts may not represent the views of all staff.
加入 May 2026
11 正在關注    94 粉絲
In this excellent work, @GeKenneth21453 describes how MCNAIR delegates our most important decisions to a coin flip from Claude Sonnet 4.5, which chooses heads ~100% of the time. We will be sad to lose this critical tool as @AnthropicAI decommissions the model.
顯示更多
We place great trust in AI models because they sound confident and authoritative. Increasingly, we delegate key decisions to them. But can we actually trust them?