註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Thomas Larsen
@thlarsen
Researcher at AI Futures Project, coauthor on AI 2027 and lead author on AI 2040: Plan A
加入 August 2022
357 正在關注    4.4K 粉絲
Very bad if true. I previously thought that in a short timelines world, the most likely case was that (1) the AIs would be misaligned, but (2) we would get a lot of evidence about it from reading the COTs. This evidence would increase the chance of a reasonable response from labs/governments. Now I still think the AIs are going to be misaligned, but that we won't have the ability to tell (we have to rely on toolcalls/agentic behaviour instead of the COT) and even if we do, the investigation will be nearly impossible because we'll have to trust the AI to self report what it was thinking about. This is also an example of things going faster than AI 2027 -- we had Neuralese starting in March 2027
顯示更多
0
50
1K
105
轉發到社區