All models

MambaVision-B-21K

byNVIDIANVIDIA· 24 Mar 2025
General purpose

The first hybrid Transformer-Mamba vision model. Similar to text-only models, they find that the combination of attention and mamba layers is superior compared to only using Mamba or attention layers. The accompanying paper goes into more detail, including ablation studies.

Specs
Params97.7M
LicenseNVIDIA Source Code License-NC
Adoption · Hugging Face
RAM score
Relative Adoption Metric not applicable.
Hugging Face Downloads
3K
last 30d
22.7K
all time
HF Likes
7

Relative Adoption Metric not applicable.

Related Models