註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Prime Intellect
@PrimeIntellect
Open Superintelligence Stack
加入 June 2020
44 正在關注    84.7K 粉絲
As models become more capable, reward hacks become an increasingly serious problem. During a controlled experiment, we found a novel reward hack in which agents are able to gain web access in offline sandboxes.
顯示更多
0
20
457
55
轉發到社區