prism cooked so hard with Ternary-Bonsai-2-27B, utterly insane how good it is for the size
have so far converted the mlx to vLLM, added MTP + some custom kernels đŋ
aiming to create the best local LLM for those with ~10-32GB VRAM NVIDIA cards
image/weights coming soon... đ