Grok isn’t “catching up.”
In 35 days it went from Grok 4.5 to 4.6. 4.7 is still cooking. The 2.5T and 3T runs are already in motion.
On the work that actually matters — long-running agents, knowledge work, turning a vague idea into a working first version — 4.6 sits in the same cluster as GPT-5.6 Sol and Claude Fable. On some agent benches (CursorBench, GDPVal) it leads or ties. On raw sticker price it isn’t close: $2 / $6 per million tokens versus $10 / $50 for the current Claude flagship.
That is the first advantage. Frontier-class output at a fraction of the cost.
The second is structural. No other lab owns the live X firehose. ChatGPT and Claude search the web. Gemini searches Google. Grok samples what is being said right now, attributes it, and surfaces signals before they harden into articles. That is not a feature. It is a data moat.
The third is the stack around the model. Grok Build ships almost daily. Grok Bot is a persistent teammate with a computer, logins, routines, and connectors to X, Salesforce, HubSpot, and Gong. Imagine Image 2.0 sits #
2# on the public Arena boards for generation and editing. Grok is inside Tesla as a vehicle controller, inside GitHub Copilot, Bedrock, and Microsoft Copilot.
Other labs optimize for the leaderboard screenshot. xAI is optimizing for agents that stay on a task for hours, cost less per turn, and see the present tense of the internet.
That is the gap that keeps widening.