注册并分享邀请链接,可获得视频播放与邀请奖励。

SemiAnalysis
@SemiAnalysis_
加入 January 2024
29 正在关注    150K 粉丝
Similar to the panic over DeepSeek R1, some uneducated people think Kimi K3’s use of linear attention (KDA) is bad for NVIDIA, HBM, DRAM, and networking because it has relatively lower KV-cache requirements. The opposite is true, and we explain why below. 👇️ 1/8🧵
显示更多
0
79
3.1K
471
转发到社区