CoolFace
Datasetpublic

ryokamoi/VisOnlyQA_eval_analysis_2

VisOnlyQA 🌐 Project Website | πŸ“„ Paper | πŸ€— Dataset | πŸ”₯ VLMEvalKit This repository contains the code and data for the paper "VisOnlyQA: Large Vision Language Models Still Struggle with Visual Perception of Geometric Information". VisOnlyQA is designed to evaluate the visual perception capability of large vision language models (LVLMs) on geometric information of scientific figures. The evaluation set includes 1,200 mlutiple choice questions in 12 visual perception tasks on… See the full description on the dataset page: https://huggingface.co/datasets/ryokamoi/VisOnlyQA_eval_analysis_2.

sourceHugging Facegpl-3.0updated 1y agoView on Hugging Face
0likes46downloads
7 commits on main
55c5e431y ago

Upload dataset

ryokamoi
e2cf7641y ago

Upload dataset

ryokamoi
b9dface1y ago

Upload dataset

ryokamoi
6ad20751y ago

Upload folder using huggingface_hub

ryokamoi
31994771y ago

Upload README.md with huggingface_hub

ryokamoi
4f659af1y ago

Upload LICENSE.md with huggingface_hub

ryokamoi
b7fdbc51y ago

initial commit

ryokamoi