ๆณจๅ†Œๅนถๅˆ†ไบซ้‚€่ฏท้“พๆŽฅ๏ผŒๅฏ่Žทๅพ—่ง†้ข‘ๆ’ญๆ”พไธŽ้‚€่ฏทๅฅ–ๅŠฑใ€‚

David Hendrickson
@TeksEdge
CEO & Founder | PhD | Startup Advisor | @Columbia | Author Generative Software Engineering | ๐Ÿ”” Follow for AI & Vibe Coding Tips ๐Ÿ‘‡
ๅŠ ๅ…ฅ July 2023
549 ๆญฃๅœจๅ…ณๆณจ    11.2K ็ฒ‰ไธ
๐Ÿ”ฅ The wildly popular DGX-Spark-friendly Qwen3.8-27B-NVFP4 now comes naked. The OrcaRouter family has already pulled 450K+ Hugging Face downloads across its FP8, GGUF, MLX, NVFP4 and BF16 builds. And now thereโ€™s a 23.4GB NVFP4 version built for NVIDIA Blackwell. ๐Ÿ‘€ ๐Ÿง  Qwen3.8-27B dense ๐Ÿšซ Abliterated / dramatically fewer refusals ๐Ÿ‘๏ธ Vision preserved ๐Ÿ› ๏ธ Tool calling preserved ๐Ÿš€ MTP preserved ๐Ÿ“š 262K context ๐Ÿ“ฆ ~23.4GB ๐Ÿ“œ Apache 2.0 And this isn't a dumb quant conversion. It mixes precision ... โšก Most FFN layers โ†’ NVFP4 ๐Ÿงฎ Attention + DeltaNet โ†’ FP8 ๐Ÿง  Final FFN layers โ†’ FP8 ๐Ÿ’พ KV cache โ†’ FP8 ๐Ÿ‘๏ธ Vision + embeddings + MTP โ†’ BF16 Target hardware ๐Ÿ”ฅ RTX 5090 ๐Ÿ”ฅ DGX Spark / GB10 ๐Ÿ”ฅ RTX PRO Blackwell ๐Ÿ”ฅ Multi-GPU RTX 50-series ๐Ÿ”ฅ B200 / B300 For AMD, Intel or Mac, there are also GGUF/MLX versions available ๐Ÿ”— HF /orcarouter/Qwen3.8-27B-Uncensored-NVFP4
ๆ˜พ็คบๆ›ดๅคš