登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

The Midas Project
@TheMidasProj
Watchdog nonprofit that monitors the practices of leading AI companies. Tracking safety updates @SafetyChanges Writing at
参加 October 2023
266 フォロー中    4.7K ファン
If you take their word for it, Grok 4.6 is now Fable-level at coding performance. You may recall that Fable was taken off the market *for weeks* due to a single jailbreak, despite a 200-page model card full of safety testing. Grok 4.6 was released (1) without a model card demonstrating any safety testing whatsoever; and (2) from a company that seems to be orders of magnitude more vulnerable to jailbreaking than its peers, with universal jailbreaks costing only ~$60 to discover. What are we even doing here? 🫩
もっと見る
1/ The AI Security Leaderboard ranks frontier AI safeguards from least to most secure. Two models tested never failed. The other two broke for under $300, after which they acted as a knowledgeable assistant for building weapons of mass destruction or hacking into computer systems.
もっと見る