Another followup on my Adversarial code review benchmarking on GPT-5.6 models. I stupidly didn't clearly call out the differences between token cost and $ cost (thanks to
@kunchenguid for pointing this out).
While tokens are imp't, $ per token are also critical that ultimately matters (tokens x price/token) even if you are using it via your ChatGPT pro subscriptions.
Revised chart shows cost in tokens, $ and time.
Luna xHigh is what I'm using for *adversarial code reviews* given the ~30% cost savings over Sol Medium.