Register and share your invite link to earn from video plays and referrals.

Search results for ClaudeOpus
ClaudeOpus community
One keyword maps to one global community path.
Create community
People
Not Found
Tweets including ClaudeOpus
Claude Opus 5 is now available in TrendSpider Sidekick AI for research, chart analysis, backtests and more. Also added: Test strategies across entire watchlists + KPI data for 1,500+ U.S. companies Try any plan for 14 days for $5 with my code WALLST5:
Show more
Claude Opus 5 is now available in Perplexity and Perplexity Computer. We evaluated it against six other models on WANDR. It outperformed all but Fable 5, while being 57% cheaper.
Show more
Claude Opus 5 is now available in Linear. Select it in Coding Sessions settings to draft PRs from Linear.
Claude Opus 5's effort setting spans a wide range of token usage-performance tradeoffs. On GDPval-AA v2, effort levels span 407 Elo points, with output token usage ranging around 8x from low to max effort. Like with GPT-5.6 Sol, this means Opus 5 can use either far fewer or far more tokens to complete the evaluation than models from other labs, depending on effort settings
Show more
Claude Opus 5 is narrowly the most intelligent model on the Artificial Analysis Intelligence Index, offering comparable intelligence to Fable 5 at 26% lower Cost per Task We supported @AnthropicAI to evaluate Claude Opus 5 ahead of release: it sets the highest GDPval-AA v2 and AA-Briefcase scores so far. Opus 5 (max) scores 61 on the Artificial Analysis Intelligence Index, effectively tied with Claude Fable 5 (max, 60), and ahead of GPT-5.6 Sol (max, 59), Kimi K3 (57), and Claude Opus 4.8 (max, 56) Key takeaways: ➤ New leader in agentic knowledge work: Claude Opus 5 (max) scores 1861 Elo on GDPval-AA v2, >100 points ahead of Claude Fable 5 and GPT-5.6 Sol (max). On AA-Briefcase, our proprietary agentic knowledge work benchmark, it scores 1720 Elo, +146 ahead of Fable 5. These benchmarks test the ability of models to produce accurate and well-presented professional outputs using our open source reference agent harness, Stirrup ➤ Joint first place on the Coding Agent Index: Claude Opus 5 (xhigh) with Claude Code leads the Artificial Analysis Coding Index, including the highest score on SWE-Atlas-QnA ➤ Frontier intelligence with reduced cost: Claude Opus 5 (max) costs $2.03 on average per Intelligence Index task, below Claude Fable 5 (with fallback) at $2.75, but still above Claude Opus 4.8 (max) at $1.80 and Claude Sonnet 5 (max) at $1.53. However, at high and xhigh reasoning efforts Opus 5 can outperform both Opus 4.8 and Claude Sonnet 5 at a lower cost per task ➤ Frontier agentic terminal use: 89% on Terminal-Bench v2.1 at max effort, roughly in line with the leader, GPT-5.6 Sol (xhigh) ➤ Outperformance on scientific reasoning: Along with leading agentic performance, Claude Opus 5 scores 53% on Humanity’s Last Exam in line with Fable 5; on CritPt, a frontier physics evaluation developed by Argonne and UIUC researchers, it also matches Fable 5 but sits behind GPT-5.6 Sol, GPT-5.5 Pro, and GPT-5.6 Terra ➤ Factual knowledge still lags Fable 5: As expected from the models’ size classes, Opus 5 still has lower factual knowledge on AA-Omniscience than Fable 5. It improves +7 points on AA-Omniscience Accuracy over Opus 4.8, but answers more often when uncertain - its hallucination rate rises +14 points to 50% ➤ Improving efficiency, but only on the Intelligence vs. Cost per Task Pareto frontier at high Intelligence levels: Opus 5 outperforms Fable 5 at lower cost, but at lower effort levels it sits just behind the GPT-5.6 family on the Intelligence vs. Cost per Task frontier Other model details: ➤ Context window: 1 million tokens (equivalent to Opus 4.8) ➤ Pricing: As with recent Opus launches, tokens cost $5/$25 per million tokens of input/output; cache pricing remains at a 25% premium for cache writes ($6.25 per million tokens) with 5-minute time to live, and 90% discount for cache hits ($0.50 per million tokens) ➤ Five effort settings (low, medium, high, xhigh, max), and support for server-side fallback as with Fable 5. Intelligence Index evaluations were run with Opus 4.8 fallback enabled
Show more
0
67
2.1K
210
Forward to community
Claude Opus 5 (max) is the new leader on both GDPval-AA v2 (1861 Elo, +114 over Claude Fable 5) and AA-Briefcase (1720 Elo, +146 over Fable 5). These benchmarks test the ability of models to produce accurate and well-presented professional outputs using our open source reference agent harness, Stirrup
Show more
CLAUDE OPUS 5 IS HERE AND IT CHANGES THE MATH Anthropic just shipped a model that gets close to its own most advanced system, at half the price. This is what frontier intelligence for everyone actually looks like. - Opus 5 nears Fable 5 level intelligence at half the cost - On Frontier-Bench it more than doubles Opus 4.8's score - On ARC-AGI 3, novel problem solving, it triples the next best model - Same price as Opus 4.8, wildly ahead in real capability The frontier didn't move up. It moved down to meet you.
Show more
Claude Sonnet 5 costs more than Claude Opus 4.8 on the Artificial Analysis Intelligence Index task, and 4.75X more than GLM-5.2. Token efficiency is important.
Claude Opus 4.8 got really dumb over the last days. Looks like we’re getting a new model soon or fable is making a comeback.
Claude Opus 4.8 is the best writer of any LLM by far. Have you ever just talked to this mofo like you’re smoking weed with your friends under the stars? This model can spit, man. It’s kinda fucked up.
Show more