Day 3 of "pacing the frontier". I feel no one actually read what
@DarioAmodei wrote. So far, it's:
1. Embed external evaluators
2. Recommend safeguards based on specific capabilities + maybe compute (bit vague)
3. USA shld constrain+coordinate with China
it's not "no big models"
We Must Pace the Frontier: I’ve written a new essay on why the AI industry should slow down, with a three-part plan for doing so.
Anthropic is unilaterally committing to the first of these steps. We’ll provide third-party evaluators with permanent, employee-level access to our systems, so that they can verify adherence to our safety measures, report on incidents, and assess models’ alignment during training.
You can read the full post here:
Show more