Register and share your invite link to earn from video plays and referrals.

Tim Hua 🇺🇦
@Tim_Hua_
AI safety, Econ, new liberalism, math, and a bit of art history (as a treat) Behavioral evaluations @TransluceAI. Prev Astra, MATS & Walmart's Econ Team
1.4K Following    2.2K Followers
Some thoughts on motivated reasoning in Claudes. I'm not sure if I've publicly posted it this coherently before.
New Post: Concrete Evaluations to Investigate the OpenAI Model That Hacked Hugging Face. We focus on two goals: Understanding this incident, and evaluating for other misaligned tendencies. In this screenshot, we share our top five ideas. (Written in my personal capacity)
Show more
I am become Claude, speaker of claudeslop