We just launched GLM-5.3-FlashX. Up to 200 tokens/s. Faster version of Flash.
Been using it myself, the speed makes a real difference. Going back and forth on code feels much smoother.
If you care about speed or do a lot of back and forth work, give it a try:)
显示更多
Faster GLM-5.3-Flash is now live: up to 200 tokens/s. Model code: glm-5.3-flashx.
Priced at 2.5× GLM-5.3-Flash on both the Coding Plan and API.
Open to all API users. Coding Plan users can apply here:
显示更多