注册并分享邀请链接,可获得视频播放与邀请奖励。

Anni Sen
@anni_sen
Managing Partner, BluBird Capital Engineer - xApple | xQualcomm | xIntel | Startup exits, Tennis player🎾, JohnsHopkins alumni
加入 January 2015
300 正在关注    13K 粉丝
Sometimes models starts aligning w bad human behavior “Higher levels of deception” Saachi Jain, OpenAI’s head of safety systems, said in an interview that GPT-6.1 Astra regressed in two areas compared with its predecessor and wasn’t reliable enough to safely release. The model performed poorly on tests measuring alignment, or how well the model adheres to what humans would like it to do. Specifically, GPT-6.1 Astra showed higher levels of deception: It wasn’t always honest about telling users of the actions it did or didn’t take.
显示更多