Register and share your invite link to earn from video plays and referrals.

SemiAnalysis
@SemiAnalysis_
Joined January 2024
29 Following    149.3K Followers
Similar to the panic over DeepSeek R1, some uneducated people think Kimi K3’s use of linear attention (KDA) is bad for NVIDIA, HBM, DRAM, and networking because it has relatively lower KV-cache requirements. The opposite is true, and we explain why below. 👇️ 1/8🧵
Show more
0
79
3.1K
471
Forward to community