Xiaomi publishes MiMo-V2.6 Pro, Flash and 9B checkpoints after live RL run
Xiaomi has turned its public MiMo-V2.6 reinforcement-learning run into downloadable Pro-RL and Flash-RL checkpoints plus a 9B Qwen distill. The flagship Pro is a sparse 1.02T-parameter model with 42B activated parameters, while Flash uses 309B total and 15B activated; both advertise 1M-token context and text, image, video and audio input under an MIT licence.