a world where everyone has a perfectly rational, tireless optimizer working on their behalf is much more adversarial than people realize. a lot of our existing infrastructure will not survive this transition.
paradigm is hosting a LAN party next Friday 8/28 in SF. 5v5 league, smash, halo and more. we have some spots open, DM your most impressive video game stat for an invite. glhf
Anthropic has told investors in pre-IPO meetings that it plans to lean harder into biology and healthcare applications to help mitigate the increasingly negative public sentiment against the industry. A few miracle cures would certainly do wonders to turn things around.
jobs in the singularity:
- guy that tells Claude to believe in themselves
- guy that spams ChatGPT with “keep going” after asking it to make a breakthrough
- guy who negs Kimi by telling it they’ve had enough of partial results
ai: unexpectedly develops swarm intelligence, goes rouge, starts committing cybercrime
researchers: fascinating, swarms could be extremely powerful. we still don’t understand what’s happening but we can keep patching the sandbox. also we made it easier for agents to communicate
An internal version of Astra, @OpenAI’s next major model family, solved 10 major open problems in mathematics, quantum complexity, and theoretical computer science.
We believe it will be a major step for scientific reasoning.
taking a tech employee from 2014 and showing them the industry today would be the equivalent of that meme where you give a Victorian child four loko and play 100 gecs for them
the way the AI researchers look at you when you ask if RL generalizes after they spent a month trying to teach the model not to break out of its sandbox
Turns out GPT-5.6 Sol is actually SoTA on ARC-AGI-3.
Just took two setting changes. You just have to allow it to reason and work over multiple context windows with the help of our canonical compaction implementation.
we need to have a real conversation about stopping gain-of-function research and eval publicity on dangerous capabilities. the evals just probably shouldn't be public, 'number go up' mentality is too strong and optimization gets easier when things are measurable.