๊ฐ€์ž… ํ›„ ์ดˆ๋Œ€ ๋งํฌ๋ฅผ ๊ณต์œ ํ•˜๋ฉด ๋™์˜์ƒ ์žฌ์ƒ ๋ฐ ์ดˆ๋Œ€ ๋ณด์ƒ์„ ๋ฐ›์„ ์ˆ˜ ์žˆ์Šต๋‹ˆ๋‹ค.

David Hendrickson
@TeksEdge
CEO & Founder | PhD | Startup Advisor | @Columbia | Author Generative Software Engineering | ๐Ÿ”” Follow for AI & Vibe Coding Tips ๐Ÿ‘‡
๊ฐ€์ž… July 2023
549 ํŒ”๋กœ์ž‰ ์ค‘    11.2K ํŒฌ
๐Ÿ”ฅ The wildly popular DGX-Spark-friendly Qwen3.8-27B-NVFP4 now comes naked. The OrcaRouter family has already pulled 450K+ Hugging Face downloads across its FP8, GGUF, MLX, NVFP4 and BF16 builds. And now thereโ€™s a 23.4GB NVFP4 version built for NVIDIA Blackwell. ๐Ÿ‘€ ๐Ÿง  Qwen3.8-27B dense ๐Ÿšซ Abliterated / dramatically fewer refusals ๐Ÿ‘๏ธ Vision preserved ๐Ÿ› ๏ธ Tool calling preserved ๐Ÿš€ MTP preserved ๐Ÿ“š 262K context ๐Ÿ“ฆ ~23.4GB ๐Ÿ“œ Apache 2.0 And this isn't a dumb quant conversion. It mixes precision ... โšก Most FFN layers โ†’ NVFP4 ๐Ÿงฎ Attention + DeltaNet โ†’ FP8 ๐Ÿง  Final FFN layers โ†’ FP8 ๐Ÿ’พ KV cache โ†’ FP8 ๐Ÿ‘๏ธ Vision + embeddings + MTP โ†’ BF16 Target hardware ๐Ÿ”ฅ RTX 5090 ๐Ÿ”ฅ DGX Spark / GB10 ๐Ÿ”ฅ RTX PRO Blackwell ๐Ÿ”ฅ Multi-GPU RTX 50-series ๐Ÿ”ฅ B200 / B300 For AMD, Intel or Mac, there are also GGUF/MLX versions available ๐Ÿ”— HF /orcarouter/Qwen3.8-27B-Uncensored-NVFP4
๋” ๋ณด๊ธฐ