注册并分享邀请链接,可获得视频播放与邀请奖励。

Joschka Braun
@BraunJoschka
AI safety researcher @ApolloResearch | Science of Scheming | prev. @MATSprogram @kasl_ai @health_nlp @uni_tue
加入 April 2020
641 正在关注    590 粉丝
I’m giving a talk on exploration hacking with Safe AI Germany (SAIGE) today at 18:00 CEST. I’ll discuss whether LLMs can learn to resist RL training, and why this matters for post-training and capability elicitation. Join here:
显示更多