注册并分享邀请链接,可获得视频播放与邀请奖励。

Stewart Slocum
@stewpervised
prev AI alignment @xai, phd @MIT
加入 September 2019
245 正在关注    1.4K 粉丝
Could existing alignment audits have caught the behaviors in the OpenAI-HuggingFace incident before they happened? As a first step, we reproduced the incident with public models. 🧵
显示更多
0
2
131
11
转发到社区