This model can fit in a 12-16gb VRAM consumer GPUs and laptops and it's almost as good as Nvidia's Qwen3.6-27b NVFP4 for agentic workflows, which is insane 🤯
@PrismML's Bonsai-27B 2-bit is performing exceptionally well!
Coding tests next.
Full eval results link: