opus 5.5, fable 5.1, and gpt-6 astra — and we're only comparing the public models here
both openai and anthropic have more advanced internal models. oai showed internal-model performance graphs around the navier-stokes work, which could give us a rough idea of how far ahead they are internally
probably around 2 months ahead?