CoolFace
Datasetpublic

xiaohei1/MMStar

MMStar (Are We on the Right Way for Evaluating Large Vision-Language Models?) 🌐 Homepage | 🤗 Dataset | 🤗 Paper | 📖 arXiv | GitHub Dataset Details As shown in the figure below, existing benchmarks lack consideration of the vision dependency of evaluation samples and potential data leakage from LLMs' and LVLMs' training data. Therefore, we introduce MMStar: an elite vision-indispensible multi-modal benchmark, aiming to ensure each curated sample… See the full description on the dataset page: https://huggingface.co/datasets/xiaohei1/MMStar.

sourceHugging Faceupdated 6mo agoView on Hugging Face
0likes38downloads
1 commits on main
4655ff96mo ago

Duplicate from Lin-Chen/MMStar

xiaohei1, Lin-Chen