注册并分享邀请链接,可获得视频播放与邀请奖励。

Liam
@liam_fallen
AI & Marketing | Ex Monday / Riverside.
加入 January 2021
678 正在关注    6K 粉丝
This was posted by the marketing team.
As models become more capable, the risks associated with developing and testing them internally also grow. We temporarily paused reinforcement learning (RL) training on our latest models intended for deployment for two weeks while we hardened and red-teamed our research environments and expanded monitoring coverage. Our largest planned frontier RL run remains on hold while smaller-scale training and evaluations validate these safeguards and establish more evidence of alignment.
显示更多