
👋 Say Hi to MiMo-Audio! Our BREAKTHROUGH in general-purpose audio intelligence. 🎯 Scaling pretraining to 100M+ hours leads to EMERGENCE of few-shot generalization across diverse audio tasks! 🔥 Post-trained MiMo-Audio-7B-Instruct: • crushes benchmarks: SOTA on MMSU, MMAU, MMAR, MMAU-Pro • outperforms Gemini-2.5-Flash on audio understanding • beats GPT-4o-Audio on complex reasoning tasks 💎 The best part? It's 100% OPEN-SOURCE Everything from tokenizer to model to evaluations! 🤗 Try it in HF Space: https://t.co/9EwFiHFdkt 📝 Tech Blog:

The full stack is open, tokenizer through , so audio-native work can be reproduced rather than only consumed through an API.
postHeads up, agent users! If you're using Xiaomi MiMo with thinking mode: When…
postMiMo-V2.5 and V2.5-Pro go open weights under MIT with day-zero SGLang and vLLM
postIntroducing MiMo-V2.5 Voice — our full-stack voice lineup for the Agent era. 🚀…Checking sign-in…
Loading comments…