BRO… GLM 5.3 Flash (ox alpha) pushed the Pareto frontier.
it matches Opus 4.8 on AA’s intelligence score but is ~45x cheaper:
> 57 intelligence vs 60 for GLM-5.3
> $0.045/task vs $0.68 for GLM-5.3
> 3x lower attention compute + 4.4x smaller KV cache at 1M context
Intelligence is getting cheaper brutally fast.
显示更多
Introducing GLM-5.3-Flash
- Leading capabilities at a highly competitive price
- Natively multimodal with a 1M-token context window
- A 320B-A18B model released under the MIT License
- Previously previewed as Ox Alpha, running entirely on Chinese AI chips
Blog:
Available now across all official platforms:
Weights:
API:
Coding Plan:
ZCode:
Chat:
AutoClaw:
显示更多