Xiaomi MiMo-V2.6 is now openโa native multimodal agent family built for large-scale reinforcement learning. ๐๐ MIT License.
๐ค
๐ MiMo-V2.6-Pro scores 46 on the Artificial Analysis Intelligence Index and reaches 71.9 on DeepSWE v1.1, 89.9 on Terminal-Bench 2.1, and 82.0 on OSWorld-Verified.
๐ง The 1.02T MoE activates 42B parameters and supports text, images, video, audio, and a 1M-token context.
โ๏ธ One mixed RL run trains coding, general, visual, and cybersecurity agents together. Pro and Flash completed 30 steps each in under six days, producing around 750K trajectories.
๐ The models support computer use, 3D creation, embodied control, coding, design, video, and music workflows.