OpenBMB

OpenBMB

9 models in Artifacts Logshuggingface.co/openbmb
MiniCPM-SALA11 Feb 2026
9BApache-2.02K · 30d

An English and Chinese 8B model with sparse attention, supporting a 1M context window.

VoxCPM-0.5B16 Sep 2025
0.5BApache-2.03.1K · 30d

A text-to-speech model with voice cloning by the OpenBMB team.

MiniCPM4.1-8B05 Sep 2025
8BApache-2.092.8K · 30d

An 8B hybrid reasoning model with 64K context, English and Chinese language support and sparse attention, released under Apache 2.0. These are explicitly marketed as on-device models, such as in their tech report, and fly under the radar because they're not challenging "frontier performance" like Qwen or DeepSeek.

MiniCPM-V-412 Jul 2025
4.1BApache-2.013.5K · 30d

A small vision model by the competent OpenBMB team. However, the license is rather restrictive, requiring attribution, disallowing model training and restricting commercial use (if >5,000 devices or >1M DAU) on top of usage restrictions.

MiniCPM4-8B06 Jun 2025
8BApache-2.036.5K · 30d

A series of small-ish models trained on 8T tokens and aimed at edge deployment. They also release CUDA Kernels for those models to use them even more efficiently.

AgentCPM-GUI13 May 2025
8BApache-2.0213 · 30d

A model trained on GUIs of smartphones to control them.

MiniCPM-o 2.612 Jan 2025
8BApache-2.0321.6K · 30d

MiniCPM is an omni-model by fusing together SigLIP, Whisper, ChatTTS, and Qwen, totaling 8B parameters. It is limited to Chinese and English audio outputs, but it is multilingual.

MiniCPM3-4B03 Sep 2024
55.2K · 30d

A very strong 4B model. They get an MMLU of 67! Crazy. We're not even there with 7B models yet at Ai2.

MiniCPM-Llama3-V-2_519 May 2024
18.2K · 30d

Two new late-fusion VLMs built on the Llama 3 8B backbone. Please reach out if you have experience with these