Register and share your invite link to earn from video plays and referrals.

Owain Evans
@OwainEvans_UK
Runs an AI Safety research group in Berkeley (Truthful AI) + Affiliate at UC Berkeley. Past: Oxford Uni, TruthfulQA, Reversal Curse. Prefer email to DM.
473 Following    20.3K Followers
New paper: LLMs should give accurate answers. 
Yet we find their answers are often biased to favor their own values and they don’t disclose this in their reasoning. 
E.g. Claude’s answer below favors Anthropic. On other tasks, Gemini & GPT-5.5 show similar biases.
Show more
Our paper on Subliminal Learning was just published in Nature! Last July we released our preprint. It showed that LLMs can transmit traits (e.g. liking owls) through data that is unrelated to that trait (numbers that appear meaningless). What’s new?🧵
Show more
0
40
888
140
Forward to community