๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

The Midas Project
@TheMidasProj
Watchdog nonprofit that monitors the practices of leading AI companies. Tracking safety updates @SafetyChanges Writing at
๊ฐ€์ž… October 2023
266 ํŒ”๋กœ์ž‰ ์ค‘    4.7K ํŒฌ
If you take their word for it, Grok 4.6 is now Fable-level at coding performance. You may recall that Fable was taken off the market *for weeks* due to a single jailbreak, despite a 200-page model card full of safety testing. Grok 4.6 was released (1) without a model card demonstrating any safety testing whatsoever; and (2) from a company that seems to be orders of magnitude more vulnerable to jailbreaking than its peers, with universal jailbreaks costing only ~$60 to discover. What are we even doing here? ๐Ÿซฉ
๋” ๋ณด๊ธฐ
1/ The AI Security Leaderboard ranks frontier AI safeguards from least to most secure. Two models tested never failed. The other two broke for under $300, after which they acted as a knowledgeable assistant for building weapons of mass destruction or hacking into computer systems.
๋” ๋ณด๊ธฐ