註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

AI Security Institute (AISI)
@AISecurityInst
We conduct scientific research to understand AI’s most serious risks and develop and test mitigations.
加入 February 2024
32 正在關注    21K 粉絲
Most AI agent evaluations boil capability down to one score. But that number hides a key choice: how much compute the agent was allowed to use. New work from our Science of Evaluation team shows why that matters. 🧵
顯示更多
0
13
414
75
轉發到社區