CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01user9000 /CLEVR-HOPE CLEVR-HOPE The CLEVR Held-Out Pair Evaluation (CLEVR-HOPE) dataset is a diagnostic dataset for testing the systematicity of VQA models. CLEVR-HOPE is a controlled setting to test whether VQA models generalize to pairs of attribute values that were not seen during either training or fine-tuning. Within CLEVR-HOPE, we refer to an unseen pair of attribute values as a Held-Out Pair (HOP). The dataset is composed of 29 sub-datasets, each for a different HOP. For each of the 29 HOPs, we… See the full description on the dataset page: https://huggingface.co/datasets/user9000/CLEVR-HOPE.imagequestion-answering10M<n<100M2 likes12k downloads1y agoHugging Face02zwcolin /clevr-multichange CLEVR-Multi-Change (30–40 objects) Two-image change-captioning data used in "Stateful Visual Encoders for Vision-Language Models" (the Multi-object Visual Differencing task). Each example is a before/after pair of a CLEVR scene with 30–40 objects and 4 simultaneous changes (add / delete / move / replace), rendered at 768×768 with a wide camera angle. Built with the CLEVR-Multi-Change engine (Johnson et al. 2017; Qiu et al. 2021). Code & paper:… See the full description on the dataset page: https://huggingface.co/datasets/zwcolin/clevr-multichange.imageimage-to-text100K<n<1M0 likes6.1k downloads4mo agoHugging Face03MMInstruction /Clevr_CoGenT_TrainA_70K_Compleximage10K<n<100K8 likes2.4k downloads2y agoHugging Face04RyanWW /Super-CLEVR Super-CLEVR: A Virtual Benchmark to Diagnose Domain Robustness in Visual Reasoning [CVPR 2023 Highlight (top 2.5%)] Paper: Super-CLEVR: A Virtual Benchmark to Diagnose Domain Robustness in Visual Reasoning Authors: Zhuowan Li, Xingrui Wang, Elias Stengel-Eskin, Adam Kortylewski, Wufei Ma, Benjamin Van Durme, Alan Yuille Dataset Description Super-CLEVR is a synthetic dataset designed to systematically study the domain robustness of visual reasoning models across… See the full description on the dataset page: https://huggingface.co/datasets/RyanWW/Super-CLEVR.imagevisual-question-answering100K<n<1M0 likes1.4k downloads3mo agoHugging Face05leonardPKU /clevr_cogen_a_trainimage10K<n<100K41 likes1.3k downloads2y agoHugging Face06MMInstruction /Clevr_CoGenT_ValAimage1K<n<10K1 likes979 downloads2y agoHugging Face07laion /clevr-webdatasetimage1M<n<10M7 likes690 downloads4y agoHugging Face08WaltonFuture /clevr-mathimage100K<n<1M1 likes425 downloads1y agoHugging Face09zechen-nlp /clevrertext100K<n<1M4 likes326 downloads2y agoHugging Face10clip-benchmark /wds_vtab-clevr_count_allimage10K<n<100K0 likes323 downloads4y agoHugging Face11nimapourjafar /mm_clevrimage10K<n<100K0 likes260 downloads2y agoHugging Face12BUAADreamer /clevr_count_70kThis dataset is borrowed from clevr_cogen_a_train image10K<n<100K3 likes260 downloads2y agoHugging Face13hunarbatra /Clevr_Complex_70kimage10K<n<100K0 likes185 downloads1y agoHugging Face14qyy752457002 /CLEVR_30K_obj_2_3image10K<n<100K0 likes155 downloads6mo agoHugging Face15sunovivid /clevr-bboximage100K<n<1M0 likes154 downloads11mo agoHugging Face16compling /CLEVR_categoriesimage100K<n<1M1 likes144 downloads1y agoHugging Face17MMInstruction /Clevr_CoGenT_TrainA_R1image10K<n<100K48 likes140 downloads2y agoHugging Face18dpdl-benchmark /clevrimage10K<n<100K2 likes139 downloads2y agoHugging Face19Aborevsky01 /CLEVR-BT-DB How to install? !pip install datasets -q from huggingface_hub import snapshot_download import pandas as pd import matplotlib.pyplot as plt # First step: download an entire datatset snapshot_download(repo_id="Aborevsky01/CLEVR-BT-DB", repo_type="dataset", local_dir='path-to-your-local-dir') # Second step: unarchive the images for VQA !unzip [path-to-your-local-dir]/[type-of-task]/images.zip # Example of the triplet (image - question -… See the full description on the dataset page: https://huggingface.co/datasets/Aborevsky01/CLEVR-BT-DB.imagevisual-question-answeringn<1K0 likes133 downloads3y agoHugging Face20clip-benchmark /wds_vtab-clevr_closest_object_distanceimage10K<n<100K1 likes132 downloads4y agoHugging Face21hunarbatra /clevr_r1 CLEVR R1 CLEVR R1 is a multimodal reasoning dataset generated with R1 to distill multimodal reasoning abilities into models. This dataset was taken from MMInstruction/Clevr_CoGenT_TrainA_R1. imagen<1K0 likes91 downloads1y agoHugging Face22Lancelot53 /clevr-changeimage1K<n<10K2 likes80 downloads1y agoHugging Face23kb049 /clevr_relimage10K<n<100K0 likes79 downloads3y agoHugging Face24dddraxxx /spatial_clevr_numbered_2000image1K<n<10K1 likes73 downloads2y agoHugging Face25MMInstruction /Clevr_CoGenT_ValBimage1K<n<10K2 likes72 downloads2y agoHugging Face26ahmedheakl /clevr-cogent-r1image10K<n<100K0 likes65 downloads1y agoHugging Face27ssampa17 /sample_clevrimagequestion-answeringn<1K1 likes63 downloads3y agoHugging Face28berhaan /clevr-tr Dataset Card for CoT Dataset Sources Repository: LLaVA-CoT GitHub Repository Paper: LLaVA-CoT on arXiv Dataset Structure Turkish - tr dataset unzip image.zip The train.jsonl file contains the question-answering data and is structured in the following format: { "id": "example_id", "image": "example_image_path", "conversations": [ {"from": "human", "value": "Lütfen resimdeki kırmızı metal nesnelerin sayısını belirtin."}, {"from": "gpt", "value":… See the full description on the dataset page: https://huggingface.co/datasets/berhaan/clevr-tr.imagevisual-question-answeringn<1K2 likes62 downloads11mo agoHugging Face29andito /clevr-math-deduplicatedimage10K<n<100K1 likes54 downloads1y agoHugging Face30dddraxxx /spatial_clevr_numbered_1000image1K<n<10K0 likes53 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.