注册并分享邀请链接,可获得视频播放与邀请奖励。

Wësche
@WescheNex1q
Day time artist and night time AI enthusiast. Building & benchmarking frontier LLMs on 4x DGX Spark clusters + Mac. Creator of Vesica Studio. Houston
加入 January 2013
637 正在关注    2.4K 粉丝
2 DGX Sparks. Qwen3.8 Flash Next NVFP4. 64 users, independent prompts + KV caches. 32,768 output tokens in 78.82s: 415.7 tok/s Video: 1K stress. Usable-context test passed 64K/user at C16, zero leakage. Testing Spark limits, not a production recipe. @NVIDIAAI
显示更多
0
9
93
10
转发到社区