CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01lmms-lab-encoder /GQA Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of GQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @inproceedings{hudson2019gqa, title={Gqa: A new dataset for real-world visual reasoning and compositional… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-encoder/GQA.image10M<n<100M34 likes39k downloads3y agoHugging Face02Voxel51 /GQA-Scene-Graph Dataset Card for GQA-35k The GQA (Visual Reasoning in the Real World) dataset is a large-scale visual question answering dataset that includes scene graph annotations for each image. This is a FiftyOne dataset with 35000 samples. Note: This is a 35,000 sample subset which does not contain questions, only the scene graph annotations as detection-level attributes. You can find the recipe notebook for creating the dataset here Installation If you haven't already… See the full description on the dataset page: https://huggingface.co/datasets/Voxel51/GQA-Scene-Graph.imageobject-detection10K<n<100K4 likes2.5k downloads2y agoHugging Face03jinyoungkim /NExT-GQA Can I Trust Your Answer? Visually Grounded Video Question Answering Introduction We study visually grounded VideoQA by forcing vision-language models (VLMs) to answer questions and simultaneously ground the relevant video moments as visual evidences. We show that this task is easy for human yet is extremely challenging for existing VLMs, revealing that the strong QA performance of these models may largely due to short-cut learning (e.g., language priors and spurious vision-text… See the full description on the dataset page: https://huggingface.co/datasets/jinyoungkim/NExT-GQA.2 likes2.4k downloads1y agoHugging Face04pppop7 /GQA Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of GQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @inproceedings{hudson2019gqa, title={Gqa: A new dataset for real-world visual reasoning and compositional question… See the full description on the dataset page: https://huggingface.co/datasets/pppop7/GQA.image10M<n<100M0 likes913 downloads9mo agoHugging Face05vikhyatk /gqaimage10K<n<100K2 likes585 downloads2y agoHugging Face06gligen /gqa_tsv0 likes498 downloads3y agoHugging Face07Mineru /GQAimage1M<n<10M1 likes485 downloads2y agoHugging Face08maelic /GQA200-coco-format GQA — General Question Answering (COCO format) This dataset is the GQA200 split of the GQA dataset (Hudson et al., 2019), reformatted in standard COCO-JSON format. GQA200 contains the top 200 object categories and 100 relations from the original GQA dataset, selected by frequency in the Stacked hybrid-attention and group collaborative learning for unbiased scene graph generation paper. This dataset has no official test split since it was used for question answering rather than… See the full description on the dataset page: https://huggingface.co/datasets/maelic/GQA200-coco-format.imageobject-detection10K<n<100K0 likes329 downloads6mo agoHugging Face09BoyangZ /GQA_llavaimage100K<n<1M0 likes312 downloads2y agoHugging Face10Feeky929 /GQA-imagesimage0 likes264 downloads5mo agoHugging Face11open-llm-leaderboard-old /details_BEE-spoke-data__smol_llama-101M-GQA Dataset Card for Evaluation run of BEE-spoke-data/smol_llama-101M-GQA Dataset Summary Dataset automatically created during the evaluation run of model BEE-spoke-data/smol_llama-101M-GQA on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_BEE-spoke-data__smol_llama-101M-GQA.0 likes235 downloads3y agoHugging Face12echarlaix /gqaGQA is a new dataset for real-world visual reasoning and compositional question answering, seeking to address key shortcomings of previous visual question answering (VQA) datasets.tabular1M<n<10M1 likes204 downloads5y agoHugging Face13wliafe /GQA200 GQA200 GQA200 is the scene-graph-generation benchmark subset of GQA. This repository uses the standard GQA200 taxonomy with 200 foreground object classes and 100 foreground predicate classes. Splits Split Images Source train 57,623 Standard GQA200 Train validation 8,209 Standard GQA200 Test The standard GQA200 Test annotations are intentionally exposed as validation. This repository does not define a separate test split. Fields… See the full description on the dataset page: https://huggingface.co/datasets/wliafe/GQA200.image10K<n<100K0 likes199 downloads2mo agoHugging Face14lance-format /gqa-testdev-balanced-lance GQA testdev-balanced (Lance Format) A Lance-formatted version of the canonical GQA testdev_balanced slice — 12,578 compositional VQA questions joined against the matching 398 images — sourced from lmms-lab/GQA. The original redistribution ships instructions and images as separate parquet configs; here they are pre-joined on image_id, so each row carries the question text, the short answer, the GQA reasoning-program tags, paired CLIP image and question embeddings, and the inline JPEG… See the full description on the dataset page: https://huggingface.co/datasets/lance-format/gqa-testdev-balanced-lance.visual-question-answering10K<n<100K1 likes197 downloads4mo agoHugging Face15dddraxxx /sealvqa_gqaimage10K<n<100K1 likes169 downloads1y agoHugging Face16BarryFutureman /GQA_for_llava_chunksimage100K<n<1M0 likes148 downloads3y agoHugging Face17open-llm-leaderboard-old /details_8Xqmff94__pooled_gqa_mix_chatml0 likes147 downloads2y agoHugging Face18DavidNguyen /GQAimage0 likes146 downloads2y agoHugging Face19open-llm-leaderboard-old /details_8xqmff94__pooled_gqa_raw0 likes145 downloads2y agoHugging Face20open-llm-leaderboard-old /details_8xqmff94__pooled_gqa_math0 likes141 downloads2y agoHugging Face21abyildirim /gqa-inpaint GQA-Inpaint Dataset GQA-Inpaint is a real image dataset to train and evaluate models for the instructional image inpainting task. Scene graphs of the GQA dataset are exploited to generate paired training data by utilizing state-of-the-art instance segmentation and inpainting methods. Dataset usage and content details are explained in the Inst-Inpaint GitHub repository. 0 likes132 downloads3y agoHugging Face22deepvk /GQA-ru GQA-ru This is a translated version of original GQA dataset and stored in format supported for lmms-eval pipeline. For this dataset, we: Translate the original one with gpt-4-turbo Filter out unsuccessful translations, i.e. where the model protection was triggered Manually validate most common errors Dataset Structure Dataset includes both train and test splits translated from original train_balanced and testdev_balanced. Train split includes 27519 images with… See the full description on the dataset page: https://huggingface.co/datasets/deepvk/GQA-ru.imagevisual-question-answering10K<n<100K7 likes127 downloads2y agoHugging Face23open-llm-leaderboard-old /details_xformAI__opt-125m-gqa-ub-6-best-for-KV-cache Dataset Card for Evaluation run of xformAI/opt-125m-gqa-ub-6-best-for-KV-cache Dataset automatically created during the evaluation run of model xformAI/opt-125m-gqa-ub-6-best-for-KV-cache on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_xformAI__opt-125m-gqa-ub-6-best-for-KV-cache.1 likes106 downloads3y agoHugging Face24open-llm-leaderboard-old /details_BEE-spoke-data__smol_llama-220M-GQA Dataset Card for Evaluation run of BEE-spoke-data/smol_llama-220M-GQA Dataset automatically created during the evaluation run of model BEE-spoke-data/smol_llama-220M-GQA on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_BEE-spoke-data__smol_llama-220M-GQA.0 likes91 downloads3y agoHugging Face25Prado2026 /gqa_answerimage10K<n<100K0 likes83 downloads5mo agoHugging Face26mm-eval /GQAimage10K<n<100K0 likes77 downloads2mo agoHugging Face27ikp23434 /gqa-tracestabular10M<n<100M0 likes76 downloads10mo agoHugging Face28alexwww94 /GQA Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of GQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @inproceedings{hudson2019gqa, title={Gqa: A new dataset for real-world visual reasoning and compositional question… See the full description on the dataset page: https://huggingface.co/datasets/alexwww94/GQA.image10M<n<100M0 likes73 downloads9mo agoHugging Face29vikhyatk /gqa-valimage10K<n<100K2 likes71 downloads2y agoHugging Face30open-llm-leaderboard-old /details_saarvajanik__facebook-opt-6.7b-gqa-ub-16-best-for-KV-cache Dataset Card for Evaluation run of saarvajanik/facebook-opt-6.7b-gqa-ub-16-best-for-KV-cache Dataset automatically created during the evaluation run of model saarvajanik/facebook-opt-6.7b-gqa-ub-16-best-for-KV-cache on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_saarvajanik__facebook-opt-6.7b-gqa-ub-16-best-for-KV-cache.2 likes69 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.