Breakthroughs emerge not from scaling or data alone, but from incorporating human feedback on preferences, a key lesson from LLMs like GPT-2 applied to world models. This preference learning is an underexplored frontier.
Here's Fabian Gura of
@odysseyml at
@raais 2026