# Xiaomi publishes MiMo V2.5 Base multimodal model

> Xiaomi published MiMo V2.5 Base, a multimodal model supporting vision, language, audio, and video understanding with long-context capabilities, hosted on Hugging Face under the MIT license.

| | |
|---|---|
| **Tool** | Xiaomi MiMo |
| **Version** | — |
| **Kind** | release |
| **Published** | 2026-04-27 |
| **Observed** | 2026-08-11 |
| **Significance** | 1/5 |
| **Breaking** | no |
| **Categories** | feature, model-support, capability |

> **Note:** this classification is below our confidence threshold and is pending human review.

## What changed


- Repository: `XiaomiMiMo/MiMo-V2.5-Base`
- Pipeline: text-generation
- Library: transformers
- Format: safetensors, fp8
- Capabilities (from tags): multimodal, vision-language, audio, video-understanding, long-context, conversational, agent
- Languages: en, zh
- License: MIT


## Sources

- [hf_model](https://huggingface.co/XiaomiMiMo/MiMo-V2.5-Base) — retrieved 2026-08-11


## Community

_No curated reactions recorded. Facts and community takes are kept in separate layers
and never blended._

---
Canonical: https://changelogs.info/xiaomi-mimo/xiaomi-publishes-mimo-v2-5-base-multimodal-model
Entity: https://changelogs.info/xiaomi-mimo
Event ID: `evt_2026-04-27_xiaomi-mimo_xiaomimimo-mimo-v2-5-base`
Licence: event synthesis © changelogs.info, CC BY 4.0. Linked sources belong to their vendors.
