๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

kaios
@kaiostephens
founder @nipuxx, @neutralityorg | data science @uwaterloo | ambassador @Alibaba_qwen
๊ฐ€์ž… June 2012
404 ํŒ”๋กœ์ž‰ ์ค‘    25K ํŒฌ
benchmarks of a 50% pruned Qwen3.6-35b-a3b and expert-specific quantization technique (made by me) 7.3gb model preforming => 51gb model, exiting to see where I can bring this technique to. I have some more things lined up too. I need a DGX spark๐Ÿ˜ญ
๋” ๋ณด๊ธฐ