Xiaomi publishes MiMo V2.5 Base multimodal model
Xiaomi published MiMo V2.5 Base, a multimodal model supporting vision, language, audio, and video understanding with long-context capabilities, hosted on Hugging Face under the MIT license.
PUBLISHED2026-04-27
OBSERVED2026-08-11
AGE4mo
SOURCES1
- Repository:
XiaomiMiMo/MiMo-V2.5-Base - Pipeline: text-generation
- Library: transformers
- Format: safetensors, fp8
- Capabilities (from tags): multimodal, vision-language, audio, video-understanding, long-context, conversational, agent
- Languages: en, zh
- License: MIT
COMMUNITY
No curated reactions recorded for this event. Facts and takes are kept in separate layers — community context is added by hand, never blended into the record above.