ryokamoi/VisOnlyQA_eval_analysis_2
VisOnlyQA π Project Website | π Paper | π€ Dataset | π₯ VLMEvalKit This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information". VisOnlyQA is designed to evaluate the visual perception capability of large vision language models (LVLMs) on geometric information of scientific figures. The evaluation set includes 1,200 mlutiple choice questions in 12 visual perception tasks onβ¦ See the full description on the dataset page: https://huggingface.co/datasets/ryokamoi/VisOnlyQA_eval_analysis_2.
Upload dataset
Upload dataset
Upload dataset
Upload folder using huggingface_hub
Upload README.md with huggingface_hub
Upload LICENSE.md with huggingface_hub
initial commit
