Register and share your invite link to earn from video plays and referrals.

Steven Adler
@sjgadler
Co-founder of Guidelight AI Standards ( ex-OpenAI safety researcher, writing at
Joined January 2018
1.1K Following    11.4K Followers
They were just trying to accurately measure the model’s abilities, which is useful for _later_ interventions, like deciding what mitigations to apply (classifiers, additional refusals training, etc) See eg: This was a capability evaluation, not a propensity evaluation
Show more