GLM-5.2 in NVFP4 is ready to serve in vLLM 🚀
@NVIDIAAI's official NVFP4 checkpoint of GLM-5.2 on Blackwell cuts the memory footprint vs FP8 while matching its accuracy across reasoning, coding, and long-context benchmarks.
Serve it today with:
vllm serve nvidia/GLM-5.2-NVFP4
🤗