3x Exited Founder/ CEO of tech cos, Chairman Emeritus- QUIN(Quad), former VC@Menlo Ventures, Author of 2 books, fmr White House fellow. All tweets personal.
Within a few years, refusing independent safety checks will cost an AI company customers.
As these systems gain the ability to write code, move money and act on our behalf, people will want more than the company’s own assurance that everything is under control.
Dario is opening the door to that scrutiny. If Anthropic lets outside evaluators inspect its work and publish what they find, customers will start asking other labs for the same access.
I think that will help responsible companies grow faster. We’re building because independent verification gives people a stronger basis for trusting AI with more consequential work.
Dario, I’d love to have you join us at Stanford Faculty Club on October 1. Let’s discuss how we make this standard practice.
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so.
Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.
You can read the full post here: