Grok 4.5 was an incredible model for the price: fast, pleasant to use, reliable, solid default model.
Grok 4.6 was a (forgivable) step in the wrong direction IMO: slower and more expensive, using way more tokens per task for a slight edge in intelligence. I get it, though. They have to climb benchmarks.
Grok 4.7 is much harder to forgive. They claimed it would be more token-efficient, and it's less by 30 to 80%. It scores worse than Grok 4.6 in various benchmarks. It's slower, it's less pleasant to use, and real-world costs come out to more than 2x above Grok 4.6, putting it over Astra's costs in real-world use.
Considering how much they've been hyping this model release up, I have to say I'm disappointed. The benchmarks don't tell the whole story, and it is pleasant to use in various real-world engineering tasks, but it feels so 2025 still.
The frontend capabilities are unacceptably bad. The 3D capabilities are nonexistent. It gets stuck in random Gemini-style loops all the time.
This was a very disappointing release. I hope that the SpaceXAI team can acknowledge that and impress us with the next one.