注册并分享邀请链接,可获得视频播放与邀请奖励。

AJ
@ItsmeAjayKV
Bullish on local AI, llm finetuning, abliteration, llama.cpp OSS contributions GPUmaxxing: 1x 3060, 1x 3090 N4 (日本語勉強中) 🎌
加入 September 2016
683 正在关注    3.5K 粉丝
I'm so happy with decode speed of Qwen3.8-Flash-Next on my 3090! It's really usable, cannot wait for next two weeks of optimizations which will push it to 30 - 40t/s. Prefill hurts thoo... 😭
显示更多
0
13
44
1
转发到社区