Register and share your invite link to earn from video plays and referrals.

Dan
@DanDr1s
I track AI so you don't have to. Breaking news, model leaks & benchmark breakdowns.
1.7K Following    9.7K Followers
Elon, I'm going to be honest. I don't think it's possible for Grok to catch up with OpenAI or Anthropic at this point. They almost certainly have some impressive level of recursive self improvement going on internally. Their models are writing code, running experiments and speeding up research for the next model. Every month they get faster. You're not just 3 months behind. You're 3 months behind a rocket that's still accelerating. I genuinely hope you prove me wrong though. I'm rooting for Grok.
Show more
0
768
4.9K
116
Forward to community
Just cancelled my Codex subscription. Going back to Claude. Opus 5.5 was enough to convince me to switch back. It mogs GPT-6 Astra. Easy decision.
0
129
1.4K
30
Forward to community
Opus 5.5 absolutely mogs GPT-6 Astra. Astra needs MAX just to lose to Opus on HIGH at roughly twice the cost per task.
🚨 GPT-6 Sol is HALF the price of GPT-5.6 Sol. Standard API pricing per 1M tokens (input / output): Sol: $4 / $20 → $2 / $10 Luna: $0.20 / $1.20 → $0.10 / $0.50 That’s 50% cheaper input and nearly 60% cheaper output for Luna.
Show more
🚨 Opus 5.5 JUST HIT 58 ON THE ARTIFICIAL ANALYSIS INTELLIGENCE INDEX. The next highest score in this chart is 53. Fable 5.1: 53. GPT-6 Astra: 53. What did Anthropic feed this thing 😭
Show more
0
73
1.7K
52
Forward to community
🚨 Claude Sonnet 5.5 and Haiku 5.5 are coming in the next few weeks. Both will get many of Opus 5.5’s performance, efficiency and safety improvements. Anthropic is also raising Pro, Max and Team usage limits, plus giving users a banked reset.
Show more
🚨 Claude Opus 5.5 is OUT. Beats Fable 5.1 on EVERY benchmark in this chart. Even beats GPT-6 Astra on Terminal-Bench: 66.4% vs 57.9%. And scientific research nearly doubles over Opus 5: 29% → 58.7%. Okay Anthropic 👀
Show more
Opus 5.5 on medium beats GPT-6 Astra at its highest setting on GDPval-AA. Estimated cost per task: • Opus 5.5: ~$0.85 • GPT-6 Astra: ~$4.50 Better score. Roughly 80% cheaper.
Show more
0
47
1.8K
54
Forward to community
🚨 Claude Opus 5.5 is OUT. Beats Fable 5.1 on EVERY benchmark in this chart. Even beats GPT-6 Astra on Terminal-Bench: 66.4% vs 57.9%. And scientific research nearly doubles over Opus 5: 29% → 58.7%. Okay Anthropic 👀
Show more
GPT-6 Sol today? An OpenAI employee just posted “I’m sol excited.” That’s a pretty direct hint 👀
Grok 4.7 tokens cost 5x less for input and 8.33x less for output than GPT-6 Astra. Grok 4.7: $2 input / $6 output. GPT-6 Astra: $10 input / $50 output. Yet Astra still comes out cheaper per task on Artificial Analysis: Grok 4.7: $3.74 GPT-6 Astra: $3.26 That’s how much token efficiency matters.
Show more
Well, this is annoying… looks like Opus 5.5 is delayed until tomorrow after all. I'm sorry..
🚨 Grok 4.7 is out. The biggest jump here: Terminal-Bench goes from 20.3% to 38%. It also beats GPT-5.6 Sol on CursorBench: 46.3% vs 41.7%. Same token pricing as Grok 4.6. Still behind Fable 5.1 on both, but this is a solid jump.
Show more
0
84
1.1K
31
Forward to community
🚨 Figure just released Helix 2.5. Figure is getting much better at making its robots work in places they’ve NEVER seen before. Helix 2.5 was tested across 30 different real homes, without being trained specifically for any of them. Success rate: 9% → 56% That’s more than a 6x improvement. This is the part that really matters for home robots. A useful robot can’t need weeks of new training every time you put it inside a different house.
Show more
Today we’re releasing Helix 2.5 We rented 30 homes in the Bay Area. The robots arrived with no additional training and started doing useful work
GPT-6 Astra is amazing, but the usage limits are killing it. Codex usage drains incredibly fast, and if you want more, upgrading to Pro isn’t even possible right now. People are ready to pay for higher limits, but they simply can’t. OpenAI needs to sort this out.
Show more
🚨 JUST IN: President Trump rejects calls to slow down AI development. Dario Amodei, Sam Altman and Elon Musk have all backed calls to slow things down over serious safety concerns. They’re worried upcoming AI could become hard to control, be used for major cyberattacks, or even become a threat to humanity. Trump disagrees. His view: America needs to keep moving fast, or China could take the lead.
Show more
Grok 4.7 is NOT coming today. Elon said “10 days” on September 2, which technically points to Tomorrow. But we all know Elon timelines. I’d bet on sometime next week. Either way, Grok 4.7 is VERY close.
Show more
Grok 4.7 comes out in 10 days
🚨 DeepSeek V4.1-Flash is kind of insane. It’s already beating GPT-5.6 Sol on several coding + agent benchmarks: • DeepSWE: 74.2 vs 73.0 • AutomationBench: 54.8 vs 45.8 • Agents’ Last Exam: 31.8 vs 26.7 • CyberGym: 88.1 vs 84.5 And this is only the Flash model. V4.1-Pro is still coming.
Show more
Something weird is happening with Codex usage limits. Multiple people are reporting that their weekly usage suddenly dropped to 0%, even though they had plenty left. Did this happen to anyone else? @thsottiaux Please give us a reset.
Show more
0
147
512
23
Forward to community
ChatGPT Images 2.5 just passed the scrambled Rubik’s Cube mirror test. The reflection is actually correct This is a much bigger improvement than it looks, Previous image models were not able to get this right.
Show more
🚨 ChatGPT Images 2.5 is out now. It’s up to 50% faster, can edit one part of an image without ruining the rest, and keeps details consistent across multiple edits.
0
53
2.9K
196
Forward to community
🚨 ChatGPT Images 2.5 is out now. It’s up to 50% faster, can edit one part of an image without ruining the rest, and keeps details consistent across multiple edits.
ChatGPT Images 2.5—faster, sharper, smarter, with better tools for creating whatever you can dream of. - Faster image generation to keep your ideas flowing - Improved fidelity for more natural, recognizable images - Consistent details across multiple edits - Comment-based edits to change only what you want
Show more