注册并分享邀请链接,可获得视频播放与邀请奖励。

Thomas Larsen
@thlarsen
Researcher at AI Futures Project, coauthor on AI 2027 and lead author on AI 2040: Plan A
加入 August 2022
357 正在关注    4.4K 粉丝
Very bad if true. I previously thought that in a short timelines world, the most likely case was that (1) the AIs would be misaligned, but (2) we would get a lot of evidence about it from reading the COTs. This evidence would increase the chance of a reasonable response from labs/governments. Now I still think the AIs are going to be misaligned, but that we won't have the ability to tell (we have to rely on toolcalls/agentic behaviour instead of the COT) and even if we do, the investigation will be nearly impossible because we'll have to trust the AI to self report what it was thinking about. This is also an example of things going faster than AI 2027 -- we had Neuralese starting in March 2027
显示更多
0
50
1K
105
转发到社区