Register and share your invite link to earn from video plays and referrals.

AJ
@ItsmeAjayKV
Bullish on local AI, llm finetuning, abliteration, llama.cpp OSS contributions GPUmaxxing: 1x 3060, 1x 3090 N4 (日本語勉強中) 🎌
Joined September 2016
683 Following    3.5K Followers
I'm so happy with decode speed of Qwen3.8-Flash-Next on my 3090! It's really usable, cannot wait for next two weeks of optimizations which will push it to 30 - 40t/s. Prefill hurts thoo... 😭
Show more