
What’s new with MiMo-V2.5 series inference? We just published a blog on our full pipeline inference optimizations for MiMo-V2.5 series, including how we pushed hybrid SWA efficiency to the limit. Read the full blog here: https://t.co/lYBEcgaVhU
Hybrid sliding-window is what makes long- and serving affordable. The writeup traces where the wins come from at each pipeline stage, useful if you serve models yourself.
postHeads up, agent users! If you're using Xiaomi MiMo with thinking mode: When…
postMiMo-V2.5 and V2.5-Pro go open weights under MIT with day-zero SGLang and vLLM
postIntroducing MiMo-V2.5 Voice — our full-stack voice lineup for the Agent era. 🚀…
postFour hours, 672 tool calls: MiMo-V2.5-Pro's long-horizon run logsChecking sign-in…
Loading comments…