註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

Inty News
@__Inty__
❤️🇺🇸 🇺🇸❤️ 我为你分享世界热点新闻 | 时政聊天群 👇
加入 August 2018
43 正在關注    752.3K 粉絲
1 token/秒 😂
You can Run DeepSeek V4 Flash (33B MoE) running on 8 GB RAM, CPU-only, zero GPU. - Cold first-token: 5.33s - Max RSS: 5.9-6.2 GiB - Process swaps: 0 The entire model stayed memory-mapped on NVMe. Linux demand-paged only the needed weights into RAM. This is not fast, It’s not production. But it proves something important: Model size vs RAM is now a performance problem, not a hard barrier. VRAM, RAM, NVMe is becoming a real memory hierarchy for local AI. -
顯示更多