注册并分享邀请链接,可获得视频播放与邀请奖励。

Anthropic
@AnthropicAI
We're an AI safety and research company that builds reliable, interpretable, and steerable AI systems. Talk to our AI assistant @claudeai on
加入 January 2021
2 正在关注    1.8M 粉丝
The fact that most individual neurons are uninterpretable presents a serious roadblock to a mechanistic understanding of language models. We demonstrate a method for decomposing groups of neurons into interpretable features with the potential to move past that roadblock.
显示更多
0
115
5.7K
980
转发到社区