登録して招待リンクを共有すると、動画再生報酬と紹介報酬を獲得できます。

John Schulman
@johnschulman2
@thinkymachines. Interested in reinforcement learning, alignment, birds, jazz music
参加 May 2021
2.2K フォロー中    82.5K ファン
Would be funny if inoculation prompting results in models that are much better at sandbox escapes and other forms of hacking because they get to spend the whole RL run practicing these things