Image-Text-to-Text
Transformers
Safetensors
English
helium1_ca
feature-extraction
custom_code

This repository contains the model weights for CASA-Helium1-VL-2B-Shared, introduced in the paper CASA: Cross-Attention over Self-Attention for Efficient Vision-Language Fusion. This is a variant of CASA-Helium1-VL-2B where the self-attention and cross-attention layers share the same parameters.

See CASA-Helium1-VL-2B for more information

Downloads last month
26
Safetensors
Model size
3B params
Tensor type
BF16
·
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Model tree for kyutai/CASA-Helium1-VL-2B-Shared

Finetuned
(1)
this model

Datasets used to train kyutai/CASA-Helium1-VL-2B-Shared

Collection including kyutai/CASA-Helium1-VL-2B-Shared

Paper for kyutai/CASA-Helium1-VL-2B-Shared