Register and share your invite link to earn from video plays and referrals.

Nicola
@NicolaManzini
Software engineer. Loves history. I made and
370 Following    4.2K Followers
After a few days @openai GPT 6 Astra reaches the top of the rankings in It beats @AnthropicAI Fable 5.1 by a sliver but is 4 times cheaper. Next on the frontier we have 2 models form @GoogleAI Gemini 3.8 and 3.7 Then even further and cheaper but still pretty good @Muse Spark and 0x Alpha GLM 5.3 Flash.
Show more
Form some early evals grok 4.6 high is better than 4.5 high in @threejs . Here some preview images. not yet released. Son on 4.5 High 4.6 High
Show more
Since Andrej @karpathy released this video I’ve been working on building a benchmark for AI coding agents on real Three.js scene generation. It’s been about a week. The site is live with: • Blind human voting on the same prompts • First quality-vs-cost rankings • 5 live scene challenges (and more coming) • Open prompt submissions Models generate full Three.js scenes from the exact same brief. Humans vote on the actual rendered results: composition, lighting, motion, whether the code works. No text scores, no cherry-picked demos. Some early surprises already: @AnthropicAI Opus 5 leads overall, but @OpenAI GPT-5.6 Luna is extremely strong for the price, and a few unspoken models are punching well above their weight. Go vote or throw hard prompts at it: P.S: Still early (static + human-only scoring for now). More models and richer evals coming. If you are interested in it comment or dm me.
Show more
@levelsio I don't understand. You change the code in prod then restart the whatever runner when the code is ready? Do you have a preview runner to check that the LLM does its job? Do you look at the code?
Show more