NarsAI/WorldVQA
WorldVQA WorldVQA: Measuring Atomic World Knowledge in Multimodal Large Language Models HomePage | Dataset | Paper | Code Abstract We introduce WorldVQA, a benchmark designed to evaluate the atomic vision-centric world knowledge of Multimodal Large Language Models (MLLMs). Current evaluations often conflate visual knowledge retrieval with reasoning. In contrast, WorldVQA decouples these capabilities to strictly measure "what the… See the full description on the dataset page: https://huggingface.co/datasets/NarsAI/WorldVQA.
067
Duplicate from moonshotai/WorldVQA
