Register and share your invite link to earn from video plays and referrals.

Dan Shipper
@danshipper
ceo @every | the only subscription you need to stay at the edge of AI
Joined January 2009
2.2K Following    141.1K Followers
BREAKING: @OpenAI just dropped GPT-6 Sol. It’s my new daily driver in Codex: not quite Astra, but close enough for much of my everyday work, faster, and 50% cheaper than 5.6 Sol. We tested it across the work we actually do at @every and ran it head to head versus Opus 5.5. Here’s my vibe check: • Writing: On a paragraph-writing task drawn from my real work, Sol scored close to Astra, which is still my top model for writing. It writes clean, minimal prose and puts the important idea first. Opus 5.5 is pleasant to work with, but its drafts still tend to bury the point. • Computer use: If you love Astra’s computer use, you’ll like Sol. An earlier Sol preview matched Astra on 17 of 18 attempts across six of our simpler Hands tasks, at a much lower token price. • Coding: Sol improves on GPT-5.6—including better Ruby code in @kieranklaassen tests—but Opus 5.5 has the higher ceiling for long, autonomous builds. One frustration: Codex’s new security classifier repeatedly stopped work we’d already authorized to ask for approval. That’s friction from the classifier, not necessarily a limitation of Sol, but it made long runs harder to leave alone. Overall: Sol feels like an S-class iPhone release: It will give you much of Astra’s power at about a fifth of Astra’s price. Opus 5.5 is the bigger surprise. In some of our tests, it matches or even beats Fable 5.1 on many tasks while also being cheaper than Opus 5. A few people on our team who were Codex converts have started to wobble with Opus 5.5. If you’re already in the Claude ecosystem you’re going to love this model. If you’re a Codex user, it’s worth a look especially for your top-end coding tasks. You’ll like Sol 6 if you spend your day in Codex reading, writing, and getting things done. It’s fast, much cheaper than Astra, and it’s the model I keep reaching for. You’ll like Opus 5.5 if you want to hand an agent a hard coding or visual project and see how far it can take it. Its best work went further than Sol’s in our tests—enough to pull some of our team back toward Claude. You can go deeper on all of our evals including each real-world task and how they're scored at the links below: Writing: Knowledge work: Reading: Full vibe checks on @every coming soon!
Show more