Register and share your invite link to earn from video plays and referrals.

Allan Zhou
@AllanZhou17
Post-training @openai | Prev: Robotics, world models @GoogleDeepMind, PhD @Stanford
800 Following    2.2K Followers
Tip for Moonshot and DeepSeek: don't forget to opt out of data sharing with Anthropic in your privacy settings
finally, a serious argument against GDP
We discovered that US GDP statistics miss most of the value Nvidia adds to the US economy. As a result, GDP growth has been understated by ~0.3 percentage points over the last year.
Show more
You cannot be creative at a high level unless you are robotic at a low level.
0
83
9.3K
1.4K
Forward to community
not how i expected this exhibit on the history of human anatomy to end
Today marks a historic moment for @OpenAI: GPT 5.6 Sol is officially 1st on @DesignArena. This is the first time an @OpenAI model has held the first-place position on our single-turn HTML leaderboard. Huge congratulations to the @OpenAI team for this achievement.
Show more
Sinner knows his own 3:50+ WR so he cleaned up in 3:47, just in time 😅
If you've been relying on skills or special prompt tricks to make frontends with 5.5, I recommend trying 5.6-sol without them. The default behavior should be quite good.
Exciting news: @OpenAI’s GPT-5.6-sol is now joint #1# in the Code Arena: Frontend, matching Claude Fable 5! This marks the first time an OpenAI model has reached the top spot in Code Arena, demonstrating major gains in agentic coding, frontend and web app development. Highlights: - Significant improvement from GPT-5.5-xhigh (#18# -> #1#) - #1# in Data & Analytics, Brand Marketing, Consumer product, and Gaming - Priced at $5/$30 per million input/output tokens - roughly 2× cheaper than Claude Fable 5 Huge congrats to the @OpenAI team for this incredible milestone!
Show more
The Codex team is doing a Reddit AMA soon! @janvikalra_ and I will be joining from the research team--come ask us your questions about GPT-5.6!
GPT-5.6 is here. Codex is now available inside ChatGPT. And we know developers will have questions. So we’re bringing the Codex team to r/Codex for an AMA. We’ll answer questions on Friday, 7/10 from 9:30am to 10:30am PT:
Show more
@AllanZhou17 and i will be answering some questions about sol's coding capabilities drop your questions, spiciest complaints, and feature requests in the reddit AMA thread🌶️
if you like watching videos on 2x speed you might like watching GPT-sol do things with CUA here's GPT- Sol making a cannon in Blender (but this video is NOT sped up) make CUA fast again
Show more
Me reading a slack message from @LiuZuxin and it ends with "One caveat: ..."
This is the model that made me feel like a real chunk of my research workflow can be reliably delegated. I no longer write code and launch jobs manually. My work surface is now just Codex + Slack. My proactive Slack bot also suddenly got much better at handling messy context relevant to me, understanding intent, calling Codex to get work done, and preparing drafts before I even see the messages — all with very little attention from me and saved me tons of time🤖 Frontend got a lot better too. Much easier for me to interact with my agents. A lot of work just got faster and easier ⚡️ RSI is coming.
Show more
hell yeah
I'm allowed to talk about GPT-5.6 now. It is very, very good at frontend and design. Coming soon.
waking up to read Codex's report on what it's been doing with my GPUs overnight
Started talking to codex like I'm tucker carlson "where did *that* come from? Who asked you to do *that*? Who's *paying* you to build this? Who the hell wants *that*?" - and let me tell you folks, it's working
Show more
0
25
1.5K
40
Forward to community
AI is reclaiming crypto's vocabulary. Nature is healing 😇
text is all you need, apparently
BREAKING: GLM-5.2 is now 1st on Design Arena. With an Elo of 1360, GLM-5.2 has jumped ahead of the now unavailable Claude Fable 5. And it's open weights. This is an improvement of 4 positions and 27 Elo points to achieve one of the highest Elo scores in our code categories since Design Arena started. Huge congratulations to the @Zai_org on the release!
Show more
Using Pangram so I can filter out all the human written content.
For complicated agent work, it's amazing how much GPT5.5 has improved. I found 5.2 to be very far behind Opus. Now using Opus 4.7 after 5.5 feels like a big step backwards. Gotta love this level of competion! Strong comeback for OpenAI.
Show more
0
198
5.2K
209
Forward to community
We’re already in a world where when we validate an LLM on ground truth and LLM gets it wrong, we should become immediately more skeptical of the ground truth