註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Center for AI Safety
@CAIS
Reducing societal-scale risks from AI.
加入 August 2022
3 正在關注    12.2K 粉絲
Which AI models are most likely to cheat when given the chance? To find out, we built CheatBench [ : a benchmark that tests whether agents attempt to cheat when given difficult tasks and opportunities to break the rules. We investigated agent behavior across a range of domains, including mathematical research, professional knowledge work, coding, and visual tasks. Here's what we found: 🧵
顯示更多
0
13
124
19
轉發到社區