注册并分享邀请链接,可获得视频播放与邀请奖励。

Unsloth AI
@UnslothAI
Run and train models locally with the Unsloth Desktop app. 🦥
加入 November 2023
478 正在关注    98.3K 粉丝
Qwen3.8-Flash can now run 1.7× faster locally with MTP!⚡️ GGUFs can reach 170 tokens/s on a RTX PRO 6000. MTP enables Qwen3.8-Flash-Next ~1.3–1.7× faster inference with no accuracy change. GGUFs: Guide:
显示更多
Qwen3.8-Flash can now be run locally! 🔥 The 125B MoE model outperforms Claude-Opus-4.6 (Max). Run on 75GB RAM via Unsloth GGUFs. Qwen3.8-Flash-Next enables CPU RAM / unified mem setups to deliver near VRAM speeds. Guide: GGUF:
显示更多
0
64
1.2K
129
转发到社区