Don't sleep on GLM. On my OOD evals, GLM has always consistently been capable where other frontier models dropped near random.
intelligence is measured as much by mastering popular evals as behaving well on unseen tasks
Introducing GLM-5.2: Frontier Intelligence, Open Weights
- Significant improvements in coding and agentic tasks
- Strong long-horizon capabilities with a 1M context window
- Two levels of reasoning effort: GLM-5.2 (max) pushes the limits, while GLM-5.2 (high) strikes a strong balance between performance and token efficiency
- MIT-licensed open weights
- Same API pricing as GLM-5.1
Tech Blog:
Weights:
API:
Coding Plan:
Chat: