Today, Xiaomi releases MiMo, our first open-source reasoning model. At 7B parameters, it’s optimized for reasoning through pre-training and post-training, surpassing OpenAI’s o1-mini and QwQ-32B-Preview on AIME 2024-2025 and LiveCodeBench v5 benchmarks. #XiaomiMiMo #MiMo7B

MiMo-7B-Base achieves significantly higher pass@k scores in reasoning benchmarks than all compared models, including the 32B baseline, across all benchmarks and evaluated k values. These results highlight the exceptional reasoning capabilities of MiMo-7B-Base.

MiMo-7B-RL not only excels in code and algorithmic tasks but also outperforms both QwQ-32B-Preview and DeepSeek-R1-Distill-Qwen-7B across general tasks, even when the reinforcement learning evaluation is limited to mathematics and code problems.

Both the MiMo-7B-Base, SFT, RL-Zero and RL model checkpoints are now open-sourced and available at: https://t.co/Odu43jXTJp For further details, please refer to the technical report: https://t.co/QAwZx5tBeD
Open checkpoints at every training stage make this a usable reference for studying how RL on math and code transfers to general reasoning at small scale.
postHeads up, agent users! If you're using Xiaomi MiMo with thinking mode: When…
postMiMo-V2.5 and V2.5-Pro go open weights under MIT with day-zero SGLang and vLLM
postIntroducing MiMo-V2.5 Voice — our full-stack voice lineup for the Agent era. 🚀…Checking sign-in…
Loading comments…