๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Noctus
@noctus91
Benchmarking for fun | Ambassador @MistralAI | Open to interesting projects & collaborations ๐Ÿ“ฉ
๊ฐ€์ž… August 2025
356 ํŒ”๋กœ์ž‰ ์ค‘    1.3K ํŒฌ
So youโ€™re telling me I can swap my local LFM2.5-2.6B from F16 to QAD Q4_0 and go from: 5.4 GB โ†’ 1.6 GB 21 โ†’ 64 tok/s 3.0s โ†’ 1.2s tool-call latency while keeping ~97% of BF16 performance? @LiquidAI what did you just do ๐Ÿ˜ญ
๋” ๋ณด๊ธฐ