註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Shauli Ravfogel
@ravfogel
Faculty fellow at NYU CDS. Previously: PhD @ BIU NLP
加入 September 2018
1.3K 正在關注    1.8K 粉絲
1/ Can LLMs introspect, i.e., reason about their internal states? Recent work claims LLMs notice when their "thoughts" get tampered with, and can report their content. We looked closely and we think it's too early to say that. Work led by @shashwat_s19 , with @tallinzen and me.
顯示更多
0
8
117
25
轉發到社區