๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

SemiAnalysis
@SemiAnalysis_
๊ฐ€์ž… January 2024
29 ํŒ”๋กœ์ž‰ ์ค‘    150K ํŒฌ
Similar to the panic over DeepSeek R1, some uneducated people think Kimi K3โ€™s use of linear attention (KDA) is bad for NVIDIA, HBM, DRAM, and networking because it has relatively lower KV-cache requirements. The opposite is true, and we explain why below. ๐Ÿ‘‡๏ธ 1/8๐Ÿงต
๋” ๋ณด๊ธฐ