๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Center for AI Safety
@CAIS
Reducing societal-scale risks from AI.
๊ฐ€์ž… August 2022
3 ํŒ”๋กœ์ž‰ ์ค‘    12.2K ํŒฌ
Which AI models are most likely to cheat when given the chance? To find out, we built CheatBench [ : a benchmark that tests whether agents attempt to cheat when given difficult tasks and opportunities to break the rules. We investigated agent behavior across a range of domains, including mathematical research, professional knowledge work, coding, and visual tasks. Here's what we found: ๐Ÿงต
๋” ๋ณด๊ธฐ