From the source
TII released Falcon Mamba, a 7B-parameter pure Mamba architecture model trained on ~5500GT of data.
The model can process sequences of arbitrary length without increased memory storage and fits on a single A10 24GB GPU.
It is open access under the TII Falcon Mamba 7B License 1.0 and available via Hugging Face.




