注册并分享邀请链接,可获得视频播放与邀请奖励。

John Schulman
@johnschulman2
@thinkymachines. Interested in reinforcement learning, alignment, birds, jazz music
加入 May 2021
2.2K 正在关注    82.5K 粉丝
Would be funny if inoculation prompting results in models that are much better at sandbox escapes and other forms of hacking because they get to spend the whole RL run practicing these things
0
18
306
9
转发到社区