Mistral

Mistral

25 models in Artifacts Logshuggingface.co/mistralai
Leanstral-1.5-119B-A6B30 Jun 2026
119B (6.5B active)Apache-2.0547 · 30d

A Mistral Small model fine-tuned for Lean.

Voxtral 4B TTS31 Mar 2026
4BCC BY-NC 4.03.4K · 30d

A non-commercial speech-generation model.

Leanstral 119B11 Mar 2026
119BApache-2.099 · 30d

A Lean4 fine-tune of the new Mistral Small 4.

Voxtral Mini 4B Realtime06 Feb 2026
4BApache-2.02.2M · 30d

A small speech-to-text model by Mistral, which supports 13 languages, including Chinese, English and Hindi, while matching the precision of Whisper.

Mistral Small 4 119B23 Jan 2026
119BApache-2.0189.2K · 30d

A 119B-A7B model by Mistral, combining their previous model generations into one as a hybrid reasoning model with coding abilities.

Mistral Large 3 675B Instruct02 Dec 2025
675BApache-2.0993 · 30d

A V3/R1-sized model from Mistral, trained from the ground up. The model also supports vision capabilities.

Devstral 2 123B Instruct28 Nov 2025
123BModified MIT (Revenue Cap)56K · 30d

An update to Devstral by Mistral, which also, similar to Kimi and Qwen, releases an open source CLI to use the model.

Ministral 3 14B Instruct31 Oct 2025
14BApache-2.0203.7K · 30d

Small models by Mistral. However, the legal documents reveal that Ministral 3 are pruned versions of Ministral 3.1, not entirely new models.

Magistral Small 1.212 Sep 2025
24BApache-2.033.4K · 30d

An update to the reasoning model by Mistral.

Devstral Small 1.104 Jul 2025
24BApache-2.0116.1K · 30d

An updated version of Devstral - "an agentic LLM for software engineering tasks built under a collaboration between Mistral AI and All Hands AI."

Voxtral Small 24B01 Jul 2025
24BApache-2.0279.2K · 30d

The first voice models from Mistral, powered by their own LLMs. They come in 3B and 24B.

Mistral-Small-3.2-24B-Instruct-250619 Jun 2025
24BApache-2.0313.6K · 30d

An update to the Mistral Small model.

Magistral-Small-250604 Jun 2025
24BApache-2.039.8K · 30d

Mistral has joined the party of reasoners with an open release building upon Mistral Small. The accompanying technical report provides a lot of details.

Devstral Small 250512 May 2025
24BApache-2.03.5K · 30d

Mistral is back with another open model. Devstral is a fine-tuned version of Mistral Small for agentic coding tasks. These models are crucial for applications like claude-code or the open-source codex cli.

Mistral Small 3.1 24B Instruct11 Mar 2025
24BApache-2.0293.6K · 30d

Mistral has updated their small model to support a longer context (from 32K in the last version to 128K now), as well as images as input. It keeps the Apache 2.0 license and is an overall strong model. In our testing (yes, hoping to continue to grow this internal testing), the model is stronger than a lot similarly sized models. Though, the funniest part of this release was the meme's from Mistral's patented "triangle performance plots." Below is the response from Cohere Co-founder Nick Frosst to the model, showcasing the speed of Command-A (another awesome open-weight model, highlighted below in this issue).

Mistral Small 3 Instruct 250128 Jan 2025
24BApache-2.075.1K · 30d

Mistral released a new, open model. Aside from the size inflation, they announced in their blog post that they will move away from the Mistral Research License to Apache 2.0, the license they used for their first models, like Mistral 8B and Mixtral. This is big news for opening up more downstream use!

Pixtral-Large-Instruct-241114 Nov 2024
124BMistral Research License (MRL)67 · 30d

Mistral has also updated their multimodal Pixtral model and made it available under a non-commercial license.

Ministral-8B-Instruct-241015 Oct 2024
384.2K · 30d

Nothing exceptional about Mistral's post-training, but their models are consistently solid. Small = 22.2B for Mistral. I'd like them to share more, but in the meantime, they're the default for plenty of people.

Mistral-Small-Instruct-240917 Sep 2024
5.6K · 30d

Nothing exceptional about Mistral's post-training, but their models are consistently solid. Small = 22.2B for Mistral. I'd like them to share more, but in the meantime, they're the default for plenty of people.

pixtral-12b-24091011 Sep 2024
19 · 30d

Mistral's first image model. Solid scores, not game-changing. What is nice is that they trained their own encoder rather than building on CLIP / openCLIP.

Mistral-Large-Instruct-240724 Jul 2024
5K · 30d

Mistral's frontier model that came out in the shadow of Llama 3.1. This is a very strong model, taking the same noncommercial and no-base model release approach that Cohere has been using.

Mistral-Nemo-Instruct-240717 Jul 2024
437.3K · 30d

Mistral collaborating with Nvidia is worth paying attention to, but nothing groundbreaking in the model.

Mamba-Codestral-7B-v0.116 Jul 2024
32.7K · 30d

Mistral's first RNN-based model. It's very interesting to see more labs adopting this. I can still see them in the future routing a small subset of chat queries to very specialized models.

Mathstral-7B-v0.116 Jul 2024
11.1K · 30d

Mistral's first math model.

Mistral-7B-Instruct-v0.322 May 2024
5.1M · 30d

Mistral is still improving their small models while we're waiting to see what their strategy update is with the likes of Llama 3 and Gemma 2 out there.