๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

Tencent Hy
@TencentHunyuan
Tencent's foundation model for text, image, video, and 3D generation.
๊ฐ€์ž… July 2024
8 ํŒ”๋กœ์ž‰ ์ค‘    51.4K ํŒฌ
We compressed Hy4-preview from 1.5TB to ๏ฝž200GiB GGUF and it still works well ! Meet MIX-STQ1_0.The trick isnโ€™t just going low, itโ€™s deciding where: calibration data picks each layerโ€™s bit-width, some down to 1.31-bit STQ1_0, some up to 2.06-bit IQ2_XXS. Same budget, lower error. Accuracy barely moves vs BF16 ๐Ÿ“Š MCP Atlas 83.7โ†’83.2 ๐Ÿ“Š SWE-Bench multi 82.9โ†’81.3 ๐Ÿ“Š MRCR 81.3โ†’81.1 ๐Ÿ“Š IFBench 73.5โ†’72.5 See the details on HF : AngelSlim/Hy4-preview-GGUF Weights & low-bit GGUFs ๐Ÿ‘‡ #LLM# #Quantization# #llamacpp# #Hy#
๋” ๋ณด๊ธฐ