註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

David Lindner
@davlindner
Making AI safer @GoogleDeepMind
加入 April 2012
347 正在關注    1.8K 粉絲
Will your AI agent secretly sabotage your work? Existing alignment evals don't directly answer this question Meet Gram: the alignment auditing tool we use to assess how likely AI agents are to engage in sabotage during internal deployments at @GoogleDeepMind
顯示更多