CoolFace
6 results

visual-spatial-reasoning

tanhuajie2001 /spatial-visual-reasoning-66kimage10K<n<100K2 likes437 downloads2y agoHugging Facetomhodemon /grounded-visual-spatial-reasoning Grounded Visual Spatial Reasoning Code for generating the annotations can be found here: github.com Dataset Summary This dataset extends the Visual Spatial Reasoning (VSR) dataset with visual grounding annotations: each caption is annotated with COCO-category object mentions, their positions , and corresponding bounding boxes in the image. Data instance Each sample instance has the following structure: Field Type Description image_file string… See the full description on the dataset page: https://huggingface.co/datasets/tomhodemon/grounded-visual-spatial-reasoning.image10K<n<100K2 likes347 downloads1y agoHugging Facejuletxara /visual-spatial-reasoningThe Visual Spatial Reasoning (VSR) corpus is a collection of caption-image pairs with true/false labels. Each caption describes the spatial relation of two individual objects in the image, and a vision-language model (VLM) needs to judge whether the caption is correctly describing the image (True) or not (False).image-classification10K<n<100K18 likes325 downloads2y agoHugging Facealbertvillanova /visual-spatial-reasoningThe Visual Spatial Reasoning (VSR) corpus is a collection of caption-image pairs with true/false labels. Each caption describes the spatial relation of two individual objects in the image, and a vision-language model (VLM) needs to judge whether the caption is correctly describing the image (True) or not (False).image-classification10K<n<100K9 likes155 downloads4y agoHugging Facetron8196 /visual-spatial-reasoning-sampled-v20 likes12 downloads1y agoHugging Face