註冊並分享邀請連結,可獲得影片播放與邀請獎勵。

klöss
@kloss_xyz
AI systems architect. prompt shaman. creative explorer. @psychanon CEO
加入 December 2021
6.2K 正在關注    76K 粉絲
fable 5.1 shipped today and the benchmarks are the least interesting part. 9 things to do with it this week: 1. reopen that expensive project: the one you killed over the bill, when the idea was fine. context-heavy jobs now run up to 45% cheaper than they did on Fable 5. 2. run your old skills through to surgically improve: no edits first, just an audit. find out which half of your instructions were babysitting a dumber model previously. 3. audit which thinking setting each workflow needs from their five levels (low/medium/high/xhigh/max): run the same job at the cheapest and the most expensive, compare. most work doesn’t need what you’re paying for. 4. hand it the bugs giving you the most trouble: a hedge fund had a crash nobody solved in four years. it took apart a vendor’s library and found it. 5. give it a real complex project and walk away. 38 hours unattended at @tryramp, and it caught its own bad result, corrected it, plus kept going. don't check if it finished, check where it wandered. 6. redo your agentic routing: one team clocked it at twice the speed of the tier below on half the tokens, which means every routing call you made in July or August could be overpriced. 7. ask for what used to get refused: the cyber filter in Claude Code now fires about 60% less per session, and the bio one fires 85% less on ordinary medical questions. delete your old workarounds. 8. build your own benchmark system, with 10 to 20 tasks. then, score every AI provider and model on it, plus every release. vendor charts are marketing. your system will be the only one that keeps these releases honest to you. 9. rebuild your content pipeline on it: @canva called the writing the standout upgrade over the code. one search company ran it blind and its judges picked the drafts 2:1 over Fable 5. if you use Claude, something you run daily is on the wrong model right now… now go find out which.
顯示更多
We’re introducing Claude Fable 5.1 and Claude Mythos 5.1. They're the world’s most advanced models for coding and knowledge work.
0
13
96
9
轉發到社區