Model personality matters in a way that benchmarks can't capture. Opus models went through a rough patch from around 4.7 to 5 where they just didn't feel "Claude-y" anymore, more like an watered-down Fable (hmmm, teacher models?). Opus 5.5 feels like working with ol' Claude again
Model complexity and the KV cache are two primary drivers for memory, especially for inference requiring increasingly large context windows.
Context windows for OpenAI’s leading models have risen 230-260X in three years to 1.05M tokens, while Meta’s Llama 4 Scout is 10X higher at 10M.
$MSFT $AMZN $META $GOOG $MU
Models across the AI ecosystem run on NVIDIA.
New models and AI labs are emerging all the time. Our CEO @JensenHuang shares the idea behind NVIDIA’s platform: build what the ecosystem needs and help everyone succeed. #AllInSummit#