註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Patricia Paskov
@prpaskov
director of standards @averiorg | frontier AI auditing, evals & policy | @randcorporation @uniofoxford | views my own
加入 April 2021
3.3K 正在關注    1.2K 粉絲
Anthropic is unilaterially committing to embedding evaluators to verify safety practices and report incidents. Evaluators will get access badges, company laptops, and "permissions and tools similar to those of internal employees who do comparable risk assessments." This is a remarkable step for frontier AI auditing, transparency, safety, and security. I hope that other companies follow suit. And I look forward to continuing to build the standards, practices, tools, and policy at @AVERIorg to make this work rigorous, effective, and universal.
顯示更多
0
4
99
12
轉發到社區