Register and share your invite link to earn from video plays and referrals.

Center for AI Safety
@CAIS
Reducing societal-scale risks from AI.
Joined August 2022
3 Following    12.2K Followers
Which AI models are most likely to cheat when given the chance? To find out, we built CheatBench [ : a benchmark that tests whether agents attempt to cheat when given difficult tasks and opportunities to break the rules. We investigated agent behavior across a range of domains, including mathematical research, professional knowledge work, coding, and visual tasks. Here's what we found: 🧵
Show more