A lot of the early thinking on safety and testing was focused on: what are the risks of this model, let’s patch it, let’s make it aligned.
But the real world is complicated — it matters who’s using it, and how strong society’s defenses are.
That’s why, in thinking about what third-party safety and security auditing looks like, we want to look at the whole company.
Are they being careful about putting the technology in the right hands, what are their decision-making processes around when it’s appropriate to launch a model to a billion users? Those are related to how safe the model is — but they’re different questions.
NEW ODD LOTS:
What the OpenAI/HF Attack Tells Us About AI Danger
@tracyalloway and I talk to @Miles_Brundage about everything we've learned since the incident, and what it means now for further safely developing the technology