Interesting that Dario thinks that pacing via inputs such as AI R&D compute is more gameable than pacing via safety evaluations/practices.
Imo it's the opposite; e.g. a requirement to spend 90% of compute on external inference (and not R&D) seems fairly hard to game.
While I am in favor of moving toward being able to pace based on safety evaluations/practices, these seem harder to define and more gameable to me; e.g. AIs or companies might game alignment evaluations, and making a judgment call on whether a safety practice is implemented appropriately seems potentially quite subjective.