Many people think that LLMs are limited by human training data, but pretraining is just establishing a common sense baseline. Most compute goes into reinforcement learning now, allowing models to move beyond the human level by exploring themselves, just as they did with Go.