註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Liam
@liam_fallen
AI & Marketing | Ex Monday / Riverside.
加入 January 2021
678 正在關注    6K 粉絲
This was posted by the marketing team.
As models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research environments and expanded monitoring coverage. Our largest planned frontier RL run remains on hold while smaller-scale training and evaluations validate these safeguards and establish more evidence of alignment.
顯示更多