註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Max Nadeau
@MaxNadeau_
Funding research to make AIs more understandable, truthful, and dependable at @coeff_giving.
加入 November 2017
612 正在關注    2K 粉絲
In light of this anecdote and the repeated pattern of models with sterling Petri scores acting misaligned in deployment, I'd like Anthropic employees to be less confident about how aligned their models just based on (current-gen) pre-deployment testing.
顯示更多
0
4
125
9
轉發到社區