
Xiaomi MiMo
Avid Artifacts readers know that Xiaomi has been working on open models for a while; its debut was exactly one year ago. The progress of its releases is remarkable, with 2.5 Pro (released under Apache 2.0) being neck and neck with other flagship models such as Kimi K2.6 and GLM-5.1 in both benchmarks and real-world usage.
Xiaomi surprised everyone by dropping a 309B-A15B MoE. The first model, which we also covered, was just a 7B dense model. Members in our subscriber-only Discord used the model and liked its writing style. However, they also found that it is lacking in terms of agentic performance and function calling.
A Qwen2.5 VL fine-tuned for robotics and autonomous driving.
A small audio model by Xiaomi, supporting auto-text, text-audio and audio-audio.
A small update to the MiMo VL model, bumping scores across the board. Similar to the previous version, the SFT version is also available.
Xiaomi, which joined the open-source community last month, has released another model. This visual model combines their previous LM with the Qwen2.5 ViT and is trained with SFT and RL.
Even Xiaomi has started releasing open models. Their first model series doesn't have to hide, following the DeepSeek playbook: Multi-token prediction, an RL-only model and an SFT->RL model. All models (Base, RL-only, SFT-only, SFT->RL) are available under the MIT license.