kyutai/CASA-Qwen2_5-VL-3B-Shared
136
This repository contains the model weights for CASA-Qwen2_5-VL-3B-Shared, introduced in the paper CASA: Cross-Attention over Self-Attention for Efficient Vision-Language Fusion. This is a variant of CASA-Qwen2_5-VL-3B where the self-attention and cross-attention layers share the same parameters.
See CASA-Qwen2_5-VL-3B for more information
