登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

David Lindner
@davlindner
Making AI safer @GoogleDeepMind
参加 April 2012
347 フォロー中    1.8K ファン
We looked at exploration hacking, a much-talked-about safety problem with so far ~no empirical work Bad news: we can make LLMs strongly resist RL elicitation Good news: we had to try pretty hard and it's easy to detect Excellent work led by @BraunJoschka @eyonjang @DamonFalck
もっと見る