Model personality matters in a way that benchmarks can't capture. Opus models went through a rough patch from around 4.7 to 5 where they just didn't feel "Claude-y" anymore, more like an watered-down Fable (hmmm, teacher models?). Opus 5.5 feels like working with ol' Claude again
Model complexity and the KV cache are two primary drivers for memory, especially for inference requiring increasingly large context windows.
Context windows for OpenAI’s leading models have risen 230-260X in three years to 1.05M tokens, while Meta’s Llama 4 Scout is 10X higher at 10M.
$MSFT $AMZN $META $GOOG $MU