tok/s used to be a vanity metric for coding models. Kimi K2.7 just made it the whole game — by outrunning you.
At ~180 tok/s (and GLM-5.2 now on every Coding Plan tier), the model generates faster than you can read, let alone review.
The bottleneck quietly moved: it's not the model anymore. It's you, sitting in the loop trying to keep up.
The instinct is to slow down and check every step. Wrong move — that just reinstalls the human as the blocker.
once the model outpaces your reading, stop reading. Hand the whole loop to the agent and put a second agent on review. Your job shifts from reading lines to designing the checks. Babysitting a 180 tok/s agent is like proofreading a printer mid-print.