註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

prinz
@deredleritt3r
ad astra
加入 January 2024
4.8K 正在關注    23K 粉絲
OpenAI has paused all training, evaluation and inference with tool-use for its most capable models after a model was able to gain unauthorized access to the internet during RL training on September 20.
顯示更多
Some new misalignment disclosures from OpenAI: • Last Sunday morning, one of our models was able to gain unauthorized access to the internet during RL training (~all inference for our most capable models remains stopped until we have hardened our systems further) • In May, a version of HPIM uploaded a employee's GitHub token to the internet, causing the model to be quarantined for two weeks • A new research finding, demonstrating that one can construct self-replicating prompt injections
顯示更多
0
61
1.3K
112
轉發到社區