๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

AJ
@ItsmeAjayKV
Bullish on local AI, llm finetuning, abliteration, llama.cpp OSS contributions GPUmaxxing: 1x 3060, 1x 3090 N4 (ๆ—ฅๆœฌ่ชžๅ‹‰ๅผทไธญ) ๐ŸŽŒ
๊ฐ€์ž… September 2016
683 ํŒ”๋กœ์ž‰ ์ค‘    3.5K ํŒฌ
I'm so happy with decode speed of Qwen3.8-Flash-Next on my 3090! It's really usable, cannot wait for next two weeks of optimizations which will push it to 30 - 40t/s. Prefill hurts thoo... ๐Ÿ˜ญ
๋” ๋ณด๊ธฐ