Register and share your invite link to earn from video plays and referrals.

Dennison
@DennisonBertram
Open for Collabs. Building a startup per week. Prev: @tallyxyz
2.1K Following    12.9K Followers
I've decided to retire SyntheticUsers. I built it because I believed agents needed an objective way to review their work from a User Persona point of view. My target was developers themselves- solo builders who couldn't identify the slop-factory that most AI generated interfaces actually are. Unfortunately I failed to get any traction on it. This is partially my fault, I didn't try hard enough for the initial user base, partially because I liked the idea more than the GTM motion of it, partially because the rapid progress in models themselves fairly quickly made me lose faith it was a problem significant enough to solve from my starting point. I'm in the process of rapid iterative exploration, so while SyntheticUsers might be retiring, I'll have many more experiments to share shortly! RIP: SyntheticUsers(dot)io
Show more
I knew it!!!!
We analyzed tens of thousands of real-world, high-reasoning Claude Opus and Fable outputs from versions 4.5 to 5 in the Text Arena and found that @claudeai's writing has changed in more ways than one: Across language complexity, patterns and markers, we now see: Claude’s answers have become much longer: - Opus 5 averages 510 words per response - This makes Opus 5 responses 3x longer than Opus 4.5’s average of 158 Responses have also become more structurally elaborate from Opus 4.5 to Opus 5: - 58% rise in average sentence length - 46% rise in clause frequency rises However, the vocabulary itself is not becoming more difficult, making Opus 5 longer and more structurally complex, but lexically simpler between 4.5 and 5: - 6.8% pt drop in long content words (46.9% to 40.1%) - 53% drop in abstract nouns (6.02 → 3.79 per 1k words) Writing habits people notice are much more common in Opus 5: - 2.3x as many em dashes - ~2x as many phrases like “load-bearing” - 50% more honesty wording like “honestly” and “frankly” Fable 5 is 38% more concise than Opus 5, averaging 316 words vs. 510. At the same time, it is nearly 2x as likely to include praise/validation or open with phrases like “yes, exactly”
Show more
The most cursed CAPTCHA I have ever seen. If you don't speak check, it says for you to "reconstruct the map by clicking and rotating the images until they align" Most of the time, they do not align.
Show more
@DennisonBertram cheers, your dapphero app helped me believe in getting shit done way back! Now it's just an addiction.
Frontier Research is the last...frontier. Something that sticks with me is @sreeramkannan's take: "The more you put into research, the more you get out of research" The opportunity is unlimited.
Introducing @YukonResearch A platform for open frontier research. Over the past 2 months, an open network of humans + AI has already: - Beat Google’s frontier quantum circuit result over 50% - Made Poolside’s open-weight model run 2.6x faster - Increased post-quantum Ethereum scaling by 3.5x - Boosted Lighter’s prover 9.5x faster Put a hard frontier science problem in front of many independent solvers, each bringing different models, harnesses, prompts and tacit knowledge. And let the best results compound. Accelerating the future of scientific progress together with AI. Bring us your hardest problem:
Show more
I've come to the conclusion that I literally hate working with the latest iteration of closed SOTA models. Infinite word salad, over-engineering, insane caution, deflection, misdirection and laziness. This is an awful way to work. Please fix it.
Show more
I just tried Devin again, I'm on the biggest plan, told it to work with cheap subagents, yet I burned my whole week in one session, plus an additional $50 in credits? What am I doing wrong? @dabit3 ?
Show more
Just had my latest Math Discovery with GPT-5.6 Sol incorporated into a real frontier math repo. I'm a certified frontier mathematician now! @jbrukh
While exciting, I'm going to need to see an hour of agentic computer use be a significant multiple of offshore outsourced talent. I agree it's coming, but thats what I think it's going to take.
An hour of agentic computer use may now be cheaper than an hour of human labor: Computer-use agent: $6-8 Offshore outsourced talent: ~$10 US talent: $30-45 "The math only gets better - inference keeps getting cheaper, and open-source models are getting good enough for a growing share of these workflows." The data on computer-use agents, from @fabrisera2000, @seema_amble, and @zephratic:
Show more
Whats the *fastest* coding model? Today, all models work great at coding when given a really good plan. I need the fastest agent. I'm tired of waiting.
How does the market value an AI harness?
Great, means the market for human generated data continues to be infinite. :-)
Today we are introducing Dyna-2, a world-action model pre-trained on one million hours of human video. At this scale, for the first time, we discovered several new scaling laws: • world-action models exhibit scaling law on human data across four orders of magnitude, from 1000 to 1,000,000 hours, • this human data scaling law implied a scaling law on never seen robot data, • both data and objective matter; world modeling and scaling on video data are essential for cross-embodiment scaling transfer to emerge 🧵
Show more
Subscription maxxing kind of sucks. It’s a miserable experience. Just let me buy a bigger subscription.
The purpose of zero or one employee companies is to remove the single largest blocker to productivity in an organization: the limited speed at which information propagates. Large organizations require levels of management solely to manage the distribution of knowledge and information. This makes them fundamentally hard to operate, slow to pivot and has a high degree of inertia. Market feedback, strategic goals, product development all take an enormous amount of time to coordinate due to the limited speed at which human employees can transmit, ingest and align around new information. Zero employee companies are fundamentally the pursuit of eliminating this bottleneck. Once you realize this, you're framing of the job to be solved totally changes.
Show more
I am fully convinced on the viability of zero-employee companies, but there are some processes that still need to be fixed. Some thoughts: 1) I don’t yet believe in all-in-one harnesses like @usenaive @polsia Cofounder, etc… not because they aren’t good ideas, but because I think it’s very hard to become “Cursor” for something as infinitely variable and diverse like running a company. 2) I don’t think this lack of faith should stop me, them, or anyone else from trying, but I haven’t quite figured out what the real niche here is. Polsia companies seem largely like garbage, and there might be a niche for that too, although given the chance I don’t think I want to build the better garbage cannon. (Although I might?) 3) running a zero employee company might actually be a skill, rather than a software stack. If it is a software stack then building the stack is mostly just planning to be acquired by Stripe if you’re lucky. This all feels very doable, but I am not yet sure the best place to start. Will share more.
Show more
In 2020 the doors of an ancient dungeon were opened by @ohjia and @wighawag. Inside was a dark and cruel evil. @EthernalWorld They managed to close the doors. Seal away the evil forever. Or so they thought. Today, I reopen the doors, and the evil returns.
Show more