I believe this issue still exists in the latest version of the Codex. After disabling subproxies for GPT-6, usage dropped significantly; a 30-minute review task on the Pro20X consumed only 1% of the usage limit.
I think I've figured out the secret of GLM5.2's huge improvements.🧐
I'm going to sleep now.
If this post gets more than 1,000 likes tomorrow, I'll write an article to reveal the reason.
Deepswe's benchmark results are my own experience.
I've used all models,
GLM 5.2 ≈ Claude Opus 4.6–4.7.
Kimi 2.7 code more like inference optimization.
Looking forward to K3.
Doubao-seed 2.1 Pro around 37% ≈ Gemini 3.5 Flash.
code are quite weak, but visual are strong.
I received festival gifts from @Alibaba_Wan at @Ali_TongyiLab thank for the thoughtful presents.
A charging adapter, a portable neck pillow, and a handheld fan.