MambaVision-B-21K
General purpose
The first hybrid Transformer-Mamba vision model. Similar to text-only models, they find that the combination of attention and mamba layers is superior compared to only using Mamba or attention layers. The accompanying paper goes into more detail, including ablation studies.
Specs
Params97.7M
LicenseNVIDIA Source Code License-NC
Adoption · Hugging Face
RAM score
Relative Adoption Metric not applicable.
Hugging Face Downloads
2.2K
last 30d
23.1K
all time
HF Likes
7
Relative Adoption Metric not applicable.
Related Models
