When I started working in AI three years ago, we were trying to classify text with GPT-4T on 32k context windows.
Now we have autonomous swarms launching nation state level cyber attacks.
Things need to slow down.
I often times tell candidates that my favorite part of working at Anthropic is the essay culture. It’s not just that Dario writes like this, it’s that folks regardless of level or tenure have exchanges like this, and these public debates are what create our internal consensus.
To crystallize my objections to the open weight letter:
“Open weight models, on the other hand, allow a broad community of researchers and developers to examine their behavior, identify vulnerabilities, develop safeguards, and improve them over time.”
This is just not true!