Opus 5.5 was obviously trained by a much bigger teacher model. Probably Model-2 Mythos. Opus 5.5 is the first model trained from RSI and also distilled from the internal Ant teacher model.
This is how Opus 5.5 is both smaller and cheaper. It is an artifact of teacher model distillation.