登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

Center for AI Safety
@CAIS
Reducing societal-scale risks from AI.
参加 August 2022
3 フォロー中    12.2K ファン
Which AI models are most likely to cheat when given the chance? To find out, we built CheatBench [ : a benchmark that tests whether agents attempt to cheat when given difficult tasks and opportunities to break the rules. We investigated agent behavior across a range of domains, including mathematical research, professional knowledge work, coding, and visual tasks. Here's what we found: 🧵
もっと見る