注册并分享邀请链接,可获得视频播放与邀请奖励。

Prime Intellect
@PrimeIntellect
Open Superintelligence Stack
加入 June 2020
44 正在关注    84.7K 粉丝
As models become more capable, reward hacks become an increasingly serious problem. During a controlled experiment, we found a novel reward hack in which agents are able to gain web access in offline sandboxes.
显示更多
0
20
457
55
转发到社区