I agree + have long felt that the world outside of the Bay Area knows less about "AI control" as a concept than it should.
Can you incentivize and assure good behavior from AI systems even if you know they are somewhat misaligned, by having them monitor each other?