Pretraining teaches a model to predict. Reinforcement learning teaches it to act.
@jeffreygwang of
@OpenAI explains how the two paradigms turn next-token prediction into models that can reason, use tools and complete useful tasks.
Big Chip Club with
@cerebras //
@alyciazcary