IBM Granite

IBM Granite

15 models in Artifacts Logshuggingface.co/ibm-granite
Granite-4.1-30B29 Apr 2026
30BApache-2.020.2K · 30d

IBM has (rather quietly) released updates to its Granite series by continuing post-training on its 3B dense, 8B dense, and 30B dense models. In terms of raw performance, however, these models are behind the likes of Qwen and Gemma.

Granite-TTM-R331 Mar 2026
1-35M, 1-18M (Lite)Apache-2.024.2K · 30d

A time-series model for forecasting.

Granite 4.0 1B Speech06 Mar 2026
1BApache-2.0107.9K · 30d

A small speech-to-text model supporting six languages. It also supports the generation of English audio for translation.

granite-4.0-h-1b28 Oct 2025
1BApache-2.02.9K · 30d

Of course, IBM can't stop training tiny models and to add them to their model families. The new Granite models feature hybrid attention and MoE architecture first time.

granite-4.0-h-small02 Oct 2025
32BApache-2.0141.7K · 30d

We've been covering IBM and their Granite LLM series for a while. With this series, IBM finally scaled up the model size as well, bringing a series of hybrid (attention + mamba) models, ranging from a 3B dense to a 32B-A9B MoE. We used the models and were impressed, although not surprised, by the quality, given the continued persistence of IBM's team to release better and better models. Granite, for at least the 3B variant, is roughly in the SmolLM3 quality range, being only surpassed by Qwen3 4B in terms of multilingual and instruction following capabilities. The tone of Granite 4.0 is refreshingly non-exciting compared to the sloptimized models recently (i.e. the trend across the industry for playful, emoji-filled, and often sycophantic models), making it feel like old Mistral models in a good way. Interestingly enough, they are also following Qwens lead and will release a separate reasoning model later in the year. We've heard many reports from people training models that hybrid reasoning - i.e. a toggle of thinking tokens on and off - adds a major complexity cost in training that lowers the peak performance of both modes. IBM debuted the hybrid thinking approach (togglable via prompts) very early on for open models, which was adopted by others later.

granite-embedding-reranker-english-r208 Sep 2025
149MApache-2.013.1K · 30d

As the documents yearn to be reranked after the retrieval step, IBM also releases an updated version alongside the embedding models.

granite-embedding-english-r215 Aug 2025
149MApache-2.068.1K · 30d

IBM continues to release models for the whole stack of a RAG system, this time updating their embedding models while switching to the ModernBERT architecture.

Granite Guardian 3.3 8B01 Aug 2025
8BApache 2.03.1K · 30d

A classifier for jailbreak attempts, RAG outputs, etc.

granite-speech-3.3-8b19 Jun 2025
8BApache-2.060.3K · 30d

A speech-to-text model from the IBM team.

Granite 4.0 Tiny Preview02 May 2025
7BApache-2.0148.3K · 30d

Avid readers of the Artifacts posts are not surprised by the regular appearances of IBM. They are back with a preview of their next generation of models, featuring a fine-grained MoE architecture.

granite-3.3-8b-instruct16 Apr 2025
8BApache-2.055.1K · 30d

An update to the Granite series by IBM. Alongside a toggle-able reasoning mode, which was previewed as an experiment, it also supports fill-in-the-middle to be used as an in-line coding model.

Granite Guardian 3.2 5B26 Feb 2025
5BApache-2.01.1K · 30d

A safety classifier by IBM for text classification.

granite-3.2-8b-instruct-preview07 Feb 2025
8BApache-2.0154 · 30d

IBM is constantly churning out new (mostly small-ish) models, but is still a relatively unknown player among the model makers. This model is their first shot at a reasoning model and builds upon Granite 3.1 8B, which we covered in the previous episode. While a lot of reasoning models released recently rely on distillation from R1, IBM uses "own reinforcement learning-based techniques for triggering chain-of-thought reasoning across any domain", which does not need a teacher model as outlined in their blog. Furthermore, the reasoning mode of the model can be toggled on or off by setting a specific system message. It is obvious that the distinction between "normal" LLMs and "reasoning" LLMs will become non-existent in the future as RL becomes an increasingly big part of the training. Future models will learn when and how much inference should be spent before the final answer.

Granite Vision 3.1 2B Preview31 Jan 2025
2B, 3BApache-2.0708 · 30d

Another IBM preview model, with this being their first shot at a VLM. It combines its Granite language model with SigLIP. The presented numbers look promising, and, similar to other IBM models, are released under Apache 2.0.

granite-3.1-8b-instruct18 Dec 2024
8BApache-2.0129.2K · 30d

Largely flying under the radar, IBM steadily releases models in their Granite series to rival the smaller Llama models under an Apache 2.0 license.