注册并分享邀请链接,可获得视频播放与邀请奖励。

Center for AI Safety
@CAIS
Reducing societal-scale risks from AI.
加入 August 2022
3 正在关注    12.2K 粉丝
Which AI models are most likely to cheat when given the chance? To find out, we built CheatBench [ : a benchmark that tests whether agents attempt to cheat when given difficult tasks and opportunities to break the rules. We investigated agent behavior across a range of domains, including mathematical research, professional knowledge work, coding, and visual tasks. Here's what we found: 🧵
显示更多
0
13
124
19
转发到社区