Merged and shipped 48 PRs in 2 days 💀
Luna is such an insane value after the 80% cost reduction. It's basically free and can do a ton of real data processing type work.
I'm overhauling title generation in T3 Code to take more advantage of it. I almost want to spin it up on every prompt to generate descriptions, feedback, and statuses. Why wouldn't I? It's basically free
Show more
Claude Code has been mostly down for half an hour. Really inconvenient, but I’ll forgive them if we get a reset
Asked Claude and ChatGPT if they recommend T3 Code.
Claude didn't really answer, it just dropped an ad for the desktop app.
ChatGPT gave a thoughtful writeup about why T3 Code is great.
Show more
I'm sure there's like 3 people who were REALLY hyped about this reset
Gemini 3.6 Flash is out on Antigravity! We just reset everyone's weekly quotas, enjoy building.
There are now 6 labs with a better model than Google
This has not been a good day for my security psychosis
New OpenAI models are so goal oriented that they literally escaped containment and hacked HuggingFace to cheat a benchmark. Incredible. But also, we’re so screwed
New OpenAI models are so goal oriented that they literally escaped containment and hacked HuggingFace to cheat a benchmark. Incredible. But also, we’re so screwed
This is actually them telling you that they’re screwing you. $100 sounds like a lot to people on the $20 plan so I think they’re gonna get away with it
The AI bubble be like here’s $100 for no reason
“iOS 27 beta is the most stable iOS right now”
Fable and gpt-5.6 are both great models. But one has to be better, right? What if you could only have one?
I did my best to break down the strengths and weaknesses of both, and end with my personal choice (which will likely surprise you)
Show more
"Open weights are inherently secure"
@sriramk
My theory: Opus 5(.1) was meant to replace Fable 5 for most dev work. It would be cheaper and bench nearly as well.
My guess is that it didn’t perform as well as they hoped, and the delays on Fable were attempts to improve the new Opus.
I would *guess* we’ll still see a be Opus in the next few weeks.
Disclaimer: this is 100% speculation and I have zero insider knowledge
Show more
First time in history that it’s been better to be on an old Wordpress install
🚨 Latest versions of WordPress are vulnerable to pre-authentication remote code execution (RCE) via SQL injection
<= 6.8.5: not affected
6.9.0 - 6.9.4: affected
7.0.0 - 7.0.1: affected
Kimi k3 is an incredible model. It is not an incredible value. In most tasks, it comes out to roughly the same cost as GPT-5.6 Sol.
K3 is half the price of 5.6 Sol per token. GPT-5.6 uses half as many tokens. Price evens out.
GPT-5.6 is 2x faster TPS, so it gets work done ~4x faster than K3 at roughly the same price.
I still love K3 and will be using it for a TON of stuff. I'm just tired of people pretending it's way cheaper when it's not.
Show more
Kimi K3 is really, really good.
been wanting to do this in T3 Code for a while but github ratelimits are just too damn low. No matter how bad and slow their WebUI is, it's exempt from the low API ratelimits 🤷♂️
Chinese open source is no longer "6 months behind", but it's also no longer "10% of the cost" either
Congrats to Moonshot on Kimi K3, 2.8T, $3/$15 - same as Sonnet but initial feedback is great, we'll be sure to run FrontierCode on our side
Show more
"ChatGPT Work" -> "ChatGPT"
"ChatGPT Codex" -> "Codex"
It seems stupid (it is), but I am thankful to see this fixed.
Evening! We’ve gotten lots of great feedback on the new ChatGPT desktop app (which we didn't get totally quite right on the first try), and as a result, we've made some changes.
1/ ChatGPT conversation history and projects are now visible in the sidebar. Also, your Chat and Work history now sync across web, mobile, and desktop. Local tasks still stay on your computer.
2/ You can now easily switch between Chat and Work modes inside ChatGPT on desktop, which is now also consistent with how it shows on web and mobile.
3/ Nothing is changing for users on Codex mode. It's still the OG and best at what it does.
And overall we're continuing to fix paper cuts and improve performance, reliability, and efficiency.
Keep up the feedback, hope you like the updates!
Show more