๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

AI Security Institute (AISI)
@AISecurityInst
We conduct scientific research to understand AIโ€™s most serious risks and develop and test mitigations.
๊ฐ€์ž… February 2024
32 ํŒ”๋กœ์ž‰ ์ค‘    21K ํŒฌ
Most AI agent evaluations boil capability down to one score. But that number hides a key choice: how much compute the agent was allowed to use. New work from our Science of Evaluation team shows why that matters. ๐Ÿงต
๋” ๋ณด๊ธฐ