So glad to see the community shipping work like GLM-5.2-Vision-NVFP4 — bolting Kimi's MoonViT onto GLM-5.2 with just a 49.5M projector and serving it on SGLang — this kind of dedication is exactly what makes the GLM ecosystem thrive 🚀
GLM-5.2 is now selectable in Claude Code via Hugging Face🤗 Inference Providers + hf-claude.
Open models are becoming easier to plug directly into real developer workflows. 😀
In just days, the community made GLM-5.2 runnable almost anywhere — GGUF for llama.cpp & Mac, MLX for Apple Silicon, NVFP4 / W4AFP8 / MXFP4 for NVIDIA & AMD, plus REAP-pruned builds for tighter footprints.
Huge thanks to everyone quantizing, pruning & testing.❤️
Find the build that fits your hardware 👇
Stopped by the @huggingface Paris office during GOSIM.
Lowkey for the open-source AI leader — hidden in the city center. People writing code wherever they feel like sitting.
🤗 keeps shipping the best open work. This vibe is half the answer.