Grok 4.5 beat Kimi K3 on the same prompt at 13x lower cost.
Very interesting experiment by AI/ML API (
@aimlapi)
Same prompt, same one-shot task
Cost per figure:
Grok 4.5 — $0.15
GPT 5.6 Sol — $0.60
Qwen 3.8 Max — $0.67
Kimi K3 — $1.98
Kimi spent ~19 minutes thinking and billed 13x more. Grok just shipped.
For production agents, cost per completed task is starting to matter a lot more than cost per token.