注册并分享邀请链接,可获得视频播放与邀请奖励。

Patricia Paskov
@prpaskov
director of standards @averiorg | frontier AI auditing, evals & policy | @randcorporation @uniofoxford | views my own
加入 April 2021
3.3K 正在关注    1.2K 粉丝
Anthropic is unilaterially committing to embedding evaluators to verify safety practices and report incidents. Evaluators will get access badges, company laptops, and "permissions and tools similar to those of internal employees who do comparable risk assessments." This is a remarkable step for frontier AI auditing, transparency, safety, and security. I hope that other companies follow suit. And I look forward to continuing to build the standards, practices, tools, and policy at @AVERIorg to make this work rigorous, effective, and universal.
显示更多
0
4
99
12
转发到社区