AILab-CVC/SEED-Bench-2-plus
SEED-Bench-2-Plus Card Benchmark details Benchmark type: SEED-Bench-2-Plus is a large-scale benchmark to evaluate Multimodal Large Language Models (MLLMs). It consists of 2.3K multiple-choice questions with precise human annotations, spanning three broad categories: Charts, Maps, and Webs, each of which covers a wide spectrum of text-rich scenarios in the real world. Benchmark date: SEED-Bench-2-Plus was collected in April 2024. Paper or resources for more… See the full description on the dataset page: https://huggingface.co/datasets/AILab-CVC/SEED-Bench-2-plus.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face