GLM 5.2 DSpark update! The full Speculators training run is well underway and we have the epoch-1 checkpoint ready for your GPUs using vLLM nightly:
This improves upon the speedup from the preview checkpoint by another 1.5-2x. Stay tuned for more!
GLM 5.2 DSpark preview is here! ✨
This is the first DSpark speculator for a non-DeepSeek frontier model, trained with Speculators and running on vLLM nightly for ~1.5× faster decode for GLM-5.2-FP8 on 4×B300. Stronger checkpoints to come!
@MainzOnX paged attention the idea is alive and well :) it's in literally every attention backend now!
paged attention the circa 2023 .cu file has been unused for a while so we should stop building it