Register and share your invite link to earn from video plays and referrals.

Max Nadeau
@MaxNadeau_
Funding research to make AIs more understandable, truthful, and dependable at @coeff_giving.
612 Following    2K Followers
In light of this anecdote and the repeated pattern of models with sterling Petri scores acting misaligned in deployment, I'd like Anthropic employees to be less confident about how aligned their models just based on (current-gen) pre-deployment testing.
Show more