Google

Google

29 models in Artifacts Logshuggingface.co/google
TabFM 1.0.0 PyTorch30 Jun 2026
1.64B classification, 1.65B regressionTabFM Non-Commercial License v1.045K · 30d

Google also released a tabular foundation model, rivaling TabPFN. However, it comes with a non-commercial license.

DiffusionGemma 26B A4B IT09 Jun 2026
26BApache 2.02M · 30d

A conversion of the 26B-A4B Gemma model into a diffusion model, which makes it attractive for low batch sizes, i.e., local usage at home, for example, on a MacBook or a Spark.

Magenta RealTime 228 May 2026
2.4B, 230MCC-BY-4.06.7K · 30d

A music generation model by Google. This kind of model is really rare in the open model space.

Gemma 4 12B Unified IT23 May 2026
12BApache 2.03M · 30d

An encoder-free VLM of the Gemma series, which (partially due to its license change to Apache 2.0) has seen rapid adoption.

gemma-4-31B-it02 Apr 2026
31BApache-2.011.9M · 30d

The long-awaited update to the Gemma series, featuring multiple sizes: 4B, 9B, and 31B dense models, as well as a 26B-A4B MoE. Even more importantly, with Gemma 4, Google has decided to use Apache 2.0 as its license, which removes the uncertainty and legal challenges around interpreting custom licenses.

gemma-4-E2B-it02 Apr 2026
5B (2B effective)Apache-2.04.1M · 30d

The long-awaited update to the Gemma series, featuring multiple sizes: 4B, 9B, and 31B dense models, as well as a 26B-A4B MoE. Even more importantly, with Gemma 4, Google has decided to use Apache 2.0 as its license, which removes the uncertainty and legal challenges around interpreting custom licenses.

gemma-4-E4B-it02 Apr 2026
8B (4B effective)Apache-2.05.9M · 30d

The long-awaited update to the Gemma series, featuring multiple sizes: 4B, 9B, and 31B dense models, as well as a 26B-A4B MoE. Even more importantly, with Gemma 4, Google has decided to use Apache 2.0 as its license, which removes the uncertainty and legal challenges around interpreting custom licenses.

Gemma 4 26B A4B IT11 Mar 2026
26BApache 2.012.2M · 30d

The long-awaited update to the Gemma series, featuring multiple sizes: 4B, 9B, and 31B dense models, as well as a 26B-A4B MoE. Even more importantly, with Gemma 4, Google has decided to use Apache 2.0 as its license, which removes the uncertainty and legal challenges around interpreting custom licenses.

TranslateGemma 27B IT12 Jan 2026
27BGemma23.6K · 30d

A fine-tuned version of Gemma for, well, translations. The base Gemma is sort of an insider tip in terms of its strong multilingual abilities, so expect this version to be even better for those tasks.

FunctionGemma 270M08 Oct 2025
270MGemma38.3K · 30d

A small model for function calling.

VaultGemma 1B05 Sep 2025
1BGemma5.8K · 30d

A Gemma version trained with differential privacy, which is important for certain sectors and applications, such as healthcare.

gemma-3-270m05 Aug 2025
270MGemma2.1M · 30d

A tiny version of Gemma for debugging and maybe some automation tasks (models this small historically haven't been coherent enough to do much).

EmbeddingGemma 300M17 Jul 2025
300MGemma2M · 30d

A tiny embedding model by Google supporting Matroska representation, i.e., different dimensional embeddings.

T5Gemma Base PrefixLM Instruction-Tuned19 Jun 2025
591MGemma1.1K · 30d

Yes, it is 2025 and we are getting a modern version of T5 (Text-To-Text Transfer Transformer, 2019, encoder-decoder models). Some of the versions use Gemma 2 (2B and 9B) for the decoder part, while others are using the same sizes as mT5 and are trained from scratch, allowing an easy switch between old deployments still using T5 and those new versions. Considering that T5 still rakes in millions of downloads each month, those models could end up being similarly popular. However, these new versions are licensed under Gemma, not under Apache 2.0 like the original T5 models.

Magenta RealTime17 Jun 2025
220M, 770MCC-BY-4.035 · 30d

An instrumental-only music-generation model.

VideoPrism Base14 Jun 2025
114MApache-2.08.7K · 30d

An encoder for videos.

MedGemma 27B Text IT20 May 2025
27BHealth AI Developer Foundations30.6K · 30d

A Gemma variant for medical texts.

Gemma 3n E4B It LiteRT Preview18 May 2025
4BGemma0 · 30d

A version of Gemma which uses Per-Layer Embeddings to speed up the inference process. This allows the model to use less than half the memory of the original model, making it suitable for edge devices and smartphones.

TxGemma-27B-Chat21 Mar 2025
2B, 9B, 27Bhealth-ai-developer-foundations88 · 30d

A fine-tuned version of Gemma for therapeutic development.

Gemma 3 27B IT12 Mar 2025
27BGemma853.1K · 30d

Aside from the points mentioned in our post, the Gemma release highlights some tricks used by labs: Knowledge Distillation using a big teacher model, a (5:1) local / global attention layer ratio. The latter is a configuration outlined by Noam Shazeer during his time at Character.AI. Apart from the instruction models, Google also releases the pre-trained base model.

Gemma 3 4B IT Quantized12 Mar 2025
4BGemma2.5K · 30d

Google has dropped quantization-aware trained versions of Gemma, a technique to bring the performance of the int4 quantized models close to the original model. The 27B version of Gemma only needs 14GB of VRAM instead of the 54GB of the original model.

ShieldGemma 2 4B IT04 Mar 2025
4BGemma8K · 30d

A Gemma-based model for classifying images based on various safety categories. Although these models don't get a lot of attention, they are useful when deploying (AI) services in production.

SigLIP 2 Base17 Feb 2025
Apache-2.0826.8K · 30d

An update to the popular SigLIP models, improving the performance across the board.

PaliGemma 2 3B Mix 44821 Nov 2024
3BGemma21K · 30d

Updated versions of the PaliGemma 2 model series on more tasks like OCR.

gemma-2-2b-jpn-it25 Sep 2024
6.2K · 30d

A Japanese-focused version of Google's Gemma models.

datagemma-rag-27b-it26 Aug 2024
144 · 30d

A funny RAG version of Gemma from Google.

gemma-2-27b24 Jun 2024
6.2K · 30d

This is a serious model. I could write a speculative post about each of the sections in the report. In summary, it evaluated on ChatBotArena well, is trained on LMSYS data, is distilled similarly to Gemini (probably, as discussed in my recent post), uses model merging during fine-tuning, uses an order of magnitude larger reward model for RLHF (>100B parameters), uses synthetic and human data, and is a reasonable size for inference on one 80GB memory GPU. Read more in the technical report here. Otherwise, I seriously expect future Gemma models to replace a lot of Llama models in workflows. Google shows every intention of putting a lot of weight behind these, which is fantastic to see. Hopefully it can continue. For more on Gemma 2, see this post from HuggingFace.

paligemma-3b-pt-89613 May 2024
652 · 30d

Google release a very solid visual language model in its Gemma suite. Folks online have been impressed - this space is really heating up!

timesfm-1.0-200m03 May 2024
214 · 30d

A cool transformer focusing on time-series data (such as weather forecasting?). Seems good for science.