註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

netrunner
@plotarmordev
Cofounder & engineer. Running local AI, building things with it, and sharing what I learn.
加入 April 2022
687 正在關注    7.2K 粉絲
On a DGX Spark, what you run still decides how fast it goes. For example, just turning on MTP in llama.cpp you can make Qwen3.8 27B on a single Spark far faster. Lots more cases like this everyday. We'll push this box to the limits until a new version is out...
顯示更多