註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

AJ
@ItsmeAjayKV
Bullish on local AI, llm finetuning, abliteration, llama.cpp OSS contributions GPUmaxxing: 1x 3060, 1x 3090 N4 (日本語勉強中) 🎌
加入 September 2016
683 正在關注    3.5K 粉絲
I'm so happy with decode speed of Qwen3.8-Flash-Next on my 3090! It's really usable, cannot wait for next two weeks of optimizations which will push it to 30 - 40t/s. Prefill hurts thoo... 😭
顯示更多
0
13
44
1
轉發到社區