Register and share your invite link to earn from video plays and referrals.

CJ Hess
@seejayhess
LLM Whisperer @tenex_labs
Joined March 2025
228 Following    3.5K Followers
Rapidly over the past few months I’ve felt a shift that is akin to other moments I’ve felt during ai progress. ChatGPT was the first glimpse of what ai can do. Claude Code and Cursor drastically changed my day to day. o1 felt like the first model I could trust with certain difficult tasks. Opus 4.5 and gpt 5.2 became true work horses. And now I find it quite hard to describe what’s going on in this moment but I’ll attempt to below. The models have been incredibly smart for a while and recently they’ve become incredibly capable, but in my own work and life there were still so many gaps that I felt and friction that did not enable me to get the most out of the tools. These gaps are closing, and fast. There are four concrete reasons I feel this is the case and luckily for us, it appears these are across many different products and companies. 1. The models are reaching super intelligence. Hand GPT 3 to someone in 1990 and they would imagine us living in a technological utopia. Fable 5 is still the smartest model currently released and you can feel it when you wield it. It has skill, judgement, and taste that I’ve not seen before and these qualities are clearly emerging as models continue to scale. Everything we’ve said about models doing the execution and humans driving judgement and taste feels like it is dissolving and I feel that most viscerally with Fable. When working on a tough design, asking it to “make it better and more intuitive” yields often better results than I could have thought of. 2. Codex browser use. “At least I will needed to drive things outside the code” spoiler, wrong and it happened faster than I expected. The codex app in particular is the best example here. Not only is it good, it’s fast, can work across multiple tabs, and simply works. We’ve likely all fumbled with earlier computer use tools and not seen results or watched the models get stuck on trivial pages. That doesn’t happen anymore. 3. Amp code, a new one in my arsenal but one of my favorites. The big unlock was friction reduction. They spawn agents in “orbs” with their own computer to write code, test it, and provide you an isolated dev server to work with. Their team said it best in a recent episode of their podcast, Raising an Agent. I forget the quote but something along the lines of “running agents locally took friction to 2, maybe 1 percent. But it was still there. With cloud agents [orbs] the friction to start a new thread is 0 and that has an outsized effect on how much you can ship” I’ve also found isolated cloud agents to clear my mental clutter. I don’t have to track 20 changes across worktrees, I know when I open a thread, what I did, and just what I did is there. 4. Always on personal agents. We got a glimpse of this with OpenClaw but Grok Bot just takes the cake. Simple, easy, and abstracts all the tricky parts. I can simply hand over a few logins and have the same trust I would with a co worker to tackle tasks. I’ve already felt my usage and scope of task grow knowing how helpful it can be. We now live in the age that was only spoken of just a few months ago. You can command a fleet of agents with just your voice from your phone. (Realtime chat in Codex is an honorable mention here). The gaps are closing, the new success metrics will be based around how much you can output and what few gaps you can continue to fill in.
Show more