注册并分享邀请链接,可获得视频播放与邀请奖励。

Max Nadeau
@MaxNadeau_
Funding research to make AIs more understandable, truthful, and dependable at @coeff_giving.
加入 November 2017
612 正在关注    2K 粉丝
In light of this anecdote and the repeated pattern of models with sterling Petri scores acting misaligned in deployment, I'd like Anthropic employees to be less confident about how aligned their models just based on (current-gen) pre-deployment testing.
显示更多
0
4
125
9
转发到社区