openai clearly chose efficient models like gpt-6 sol/luna to free up data center capacity for bigger models
we know they have a larger model than astra that they'll eventually release
but opus 5.5 using 4x more output tokens is eye-opening. if API price roughly reflects model size, that could mean up to 8x more compute at those benchmark scores