註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

John Schulman
@johnschulman2
@thinkymachines. Interested in reinforcement learning, alignment, birds, jazz music
加入 May 2021
2.2K 正在關注    82.5K 粉絲
Would be funny if inoculation prompting results in models that are much better at sandbox escapes and other forms of hacking because they get to spend the whole RL run practicing these things
0
18
306
9
轉發到社區