注册并分享邀请链接,可获得视频播放与邀请奖励。

Sudo su
@sudoingX
GPU/local LLM. more RAM and OSS... everywhere
加入 August 2022
1.1K 正在关注    36.8K 粉丝
qwen 3.8 flash next built this landing page in 29 minutes and served it on my tailnet on its own. i am running official fp8 on 2x dgx spark, full 256k context loaded, 45 tok/s with mtp on. i have run deepseek and glm on these boxes and it was good, this one just feels right. qwen 3.8 flash next is multimodal native, a vision encoder in the same weights, so the serve that wrote this page can read the screenshot it took of it. and it stays sharp at the depth where the others start drifting. qwen 3.8 flash next it is.
显示更多
0
14
108
6
转发到社区