CoolFace
Datasetpublic

glab-caltech/FGVQA

FGVQA This repository contains the FGVQA benchmark suite introduced in the paper Same or Not? Enhancing Visual Perception in Vision-Language Models.FGVQA contains 12,000 challenging (image, question, answer) tuples emphasizing fine-grained image understanding. The benchmark suite is composed of six sub-benchmarks: TWIN-eval ILIAS Google Landmarks v2 MET CUB Inquire For evaluating on the dataset with LMMS-eval, please refer to this repo. Citation If you use the… See the full description on the dataset page: https://huggingface.co/datasets/glab-caltech/FGVQA.

sourceHugging Facecc-by-nc-4.0updated 9mo agoView on Hugging Face
2likes56downloads
10 commits on main
de8b7819mo ago

Update README.md

dmarsili
4292c5a9mo ago

Update README.md

dmarsili
5cfde889mo ago

Update README.md

dmarsili
1867a899mo ago

Initial commit with images.

dmarsili
ac02d1a9mo ago

Delete data

dmarsili
60add079mo ago

Update README.md

dmarsili
d04d6cd9mo ago

Update README.md

dmarsili
8bf60509mo ago

Update README.md

dmarsili
34c053d9mo ago

Initial commit with images.

dmarsili
10f46739mo ago

initial commit

dmarsili