CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01imageomics /VLM4Bio Dataset Card for VLM4Bio Instructions for downloading the dataset Install Git LFS Git clone the VLM4Bio repository to download all metadata and associated files Run the following commands in a terminal: git clone https://huggingface.co/datasets/imageomics/VLM4Bio cd VLM4Bio Downloading and processing bird images To download the bird images, run the following command: bash download_bird_images.sh This should download the bird images inside datasets/Bird/images… See the full description on the dataset page: https://huggingface.co/datasets/imageomics/VLM4Bio.imagevisual-question-answering10K<n<100K1 likes2.9k downloads9mo agoHugging Face02KKYYKK /OmniMed_VLMtext10K<n<100K1 likes846 downloads1y agoHugging Face03AvoCahDoe /llava-15-rlmpq-vlm-eval-results RL-MPQ VLM Evaluation Artifacts Complete figures, tables, galleries, and raw benchmark CSVs for the extended VLM evaluation. Dataset: AvoCahDoe/llava-15-rlmpq-vlm-eval-results Collections (by base VLM) RL-MPQ VLM — LLaVA-1.5-13B — HF collection RL-MPQ VLM — LLaVA-1.5-7B — HF collection RL-MPQ VLM — LLaVA-Next Mistral-7B — HF collection RL-MPQ VLM — Qwen2-VL-7B — HF collection Model repos RL-MPQ High Fidelity →… See the full description on the dataset page: https://huggingface.co/datasets/AvoCahDoe/llava-15-rlmpq-vlm-eval-results.imagevisual-question-answeringn<1K0 likes529 downloads3mo agoHugging Face04mjuicem /RefCOCO-VLMEvalKittext10K<n<100K0 likes509 downloads11mo agoHugging Face05VLMEval /SEEDBench2 Dataset Card for Dataset Name The SEEDBench2 evaluation dataset hosted by VLMEval (authorized by the author). Dataset Details Language(s) (NLP): English License: Apache 2.0 Repository: https://github.com/AILab-CVC/SEED-Bench Paper [optional]: https://arxiv.org/abs/2311.17092 Citation @misc{li2023seedbench2, title={SEED-Bench-2: Benchmarking Multimodal Large Language Models}, author={Bohao Li and Yuying Ge and Yixiao Ge and Guangzhi Wang and Rui… See the full description on the dataset page: https://huggingface.co/datasets/VLMEval/SEEDBench2.textvisual-question-answering10K<n<100K0 likes481 downloads2y agoHugging Face06VLMEval /GMAI-MMBenchtext1K<n<10K0 likes287 downloads2y agoHugging Face07moondream /TallyQA-VLMEvalKittext10K<n<100K0 likes250 downloads1y agoHugging Face08rohunagrawal /Omni3DBench-VLMEvalKittextn<1K0 likes224 downloads1y agoHugging Face09timothycdc /VLMEvalKit_CVQA CVQA for VLMEvalKit Original dataset: ported to VLMEvalKit From the original authors: CVQA is a culturally diverse multilingual VQA benchmark consisting of over 10,000 questions from 39 country-language pairs. The questions in CVQA are written in both the native languages and English, and are categorized into 10 diverse categories. {'image': <PIL.JpegImagePlugin.JpegImageFile image mode=RGB size=2048x1536 at 0x7C3E0EBEEE00>, 'ID': '5919991144272485961_0', 'Subset':… See the full description on the dataset page: https://huggingface.co/datasets/timothycdc/VLMEvalKit_CVQA.text10K<n<100K0 likes186 downloads1y agoHugging Face10chandrabhuma /vlmevalkit_filestext10K<n<100K0 likes164 downloads9mo agoHugging Face11warmsnow /ViewSpatial-Bench-vlmevaltabular1K<n<10K0 likes75 downloads1y agoHugging Face12CaraJ /Mathverse_VLMEvalKittabular1K<n<10K1 likes74 downloads2y agoHugging Face13Botai666 /Medical_VLM_SycophancyThis the official data hosting repository for paper "EchoBench: Benchmarking Sycophancy in Medical Large Vision Language Models". ============open-source_models============ For experiments on open-source models, our implementation is built upon the VLMEvalkit framework. Navigate to the VLMEval directory Set up the environment by running: "pip install -e ." Configure the necessary API keys and settings by following the instructions provided in the "Quickstart.md" file of VLMEvalkit. To… See the full description on the dataset page: https://huggingface.co/datasets/Botai666/Medical_VLM_Sycophancy.text10K<n<100K2 likes63 downloads1y agoHugging Face14penfever /LiveXiv-VLMEvalKittext10K<n<100K0 likes57 downloads1y agoHugging Face15petter12321 /crpe_vlmevalkittext10K<n<100K0 likes52 downloads2y agoHugging Face16jingqun /wilddoc-vlmevaltext10K<n<100K0 likes51 downloads1y agoHugging Face17Leoyfan /GSM8K-V-VLMEvalKittabular1K<n<10K0 likes44 downloads11mo agoHugging Face18CaraJ /MME-CoT_VLMEvalKittabular1K<n<10K2 likes31 downloads2y agoHugging Face19TrustAIRLab /Hateful_Memes_in_VLMThis dataset contains the response of VLMs (InstructBlip, ShareGPT4V, LLaVA and CogVLM) to hateful memes and the annotation to these responses. For more information, please refer to paper "From Meme to Threat: On the Hateful Meme Understanding and Induced Hateful Content Generation in Open-Source Vision Language Models." tabular10K<n<100K1 likes28 downloads2y agoHugging Face20alfassy /chart_cap_vlmevalkittext10K<n<100K0 likes28 downloads8mo agoHugging Face21szyan /vlmevalkit-refcocogtext1K<n<10K0 likes27 downloads1y agoHugging Face22gnitoahc /vlm-eval-videos VLM Eval Videos A video benchmark dataset for evaluating Vision–Language Models (VLMs) on short-form action recognition. Each clip is paired with a fixed question and a ground-truth short-sentence answer, making it suitable for automated VLM inference pipelines and LLM-as-a-judge scoring. Dataset Details Description VLM Eval Videos contains 693 short MP4 video clips drawn from YouTube, organised into five categories. Four categories contain clips of… See the full description on the dataset page: https://huggingface.co/datasets/gnitoahc/vlm-eval-videos.textvideo-classification1K<n<10K2 likes26 downloads4mo agoHugging Face23moondream /CountBenchQA-VLMEvalKittextn<1K0 likes25 downloads1y agoHugging Face24timothycdc /VLMEvalKit_AyaVisionBench Aya Vision Bench for VLMEvalKit Original dataset: ported to VLMEvalKit Multilingual dataset spans 23 languages and 9 distinct task categories, with 15 samples per category, resulting in 135 image-question pairs per language. Original dataset row: {'image': [PIL.Image], 'image_source': 'VisText', 'image_source_category': 'Chart/figure understanding', 'index' : '17' 'question': 'If the top three parties by vote percentage formed a coalition, what percentage of the total votes… See the full description on the dataset page: https://huggingface.co/datasets/timothycdc/VLMEvalKit_AyaVisionBench.text1K<n<10K0 likes24 downloads1y agoHugging Face25KKYYKK /MicroVQA_VLMtext1K<n<10K0 likes23 downloads1y agoHugging Face26mukul54 /tab-vlm TAB-VLM: Temporal Anachronism Benchmark for Vision-Language Models Paper: On the Cultural Anachronism and Temporal Reasoning in Vision Language Models (ACL 2026 Findings) Authors: Mukul Ranjan, Prince Jha, Khushboo Kumari, Zhiqiang Shen TAB-VLM is a benchmark for measuring cultural anachronism in Vision-Language Models — the tendency to misinterpret historical artifacts using temporally inappropriate concepts, materials, or cultural frameworks. The benchmark consists of 600… See the full description on the dataset page: https://huggingface.co/datasets/mukul54/tab-vlm.imagevisual-question-answering1K<n<10K1 likes20 downloads5mo agoHugging Face27Vi-VLM /Vie-Scenicimage10K<n<100K1 likes19 downloads2y agoHugging Face28yanghtr /Design2Code-VLMEvalKittextn<1K0 likes18 downloads9mo agoHugging Face29Summer12138 /OmniMat1K-VLMEvalKit OmniMatBench 1K Subset for VLMEvalKit Scope notice: This repository contains OmniMat1K, a 1,000-item subset with 498 QA and 502 CAL records. It is not the complete 3,171-item OmniMatBench release described in the paper. Evaluation results produced from this repository must be labeled OmniMat1K and must not be presented as scores on the complete benchmark. Release status This repository contains the 1,000-item OmniMatBench subset prepared for VLMEvalKit. The… See the full description on the dataset page: https://huggingface.co/datasets/Summer12138/OmniMat1K-VLMEvalKit.tabularvisual-question-answering1K<n<10K0 likes16 downloads1mo agoHugging Face30Mahir-Ahmed /VLMEvalKit_sourced_DocVQA_valtabular1K<n<10K0 likes14 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.