I've heard some people say that Composer 2 is "just Kimi 2.5", which is easy to disprove if you look at any benchmarks... or, huh, if you run the model for like 30 seconds. Anyway: this paper shows the humongous training effort that went behind our model. Please check it out.