CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01lmms-eval /Video-MMEtext1K<n<10K96 likes40k downloads2y agoHugging Face02lmms-lab-encoder /LMMs-Eval-Liteimage1K<n<10K7 likes6.7k downloads2y agoHugging Face03lmms-eval /egoschematext10K<n<100K9 likes5.9k downloads2y agoHugging Face04lmms-eval /LVBenchtext1K<n<10K6 likes5k downloads1y agoHugging Face05lmms-eval /NExTQAtabular10K<n<100K6 likes3.8k downloads2y agoHugging Face06lmms-eval /YouCook2text1K<n<10K3 likes3k downloads2y agoHugging Face07lmms-eval /LiveBenchhttps://arxiv.org/abs/2407.12772 image1K<n<10K5 likes2.9k downloads2y agoHugging Face08lmms-eval /TempCompasstext1K<n<10K6 likes2.4k downloads2y agoHugging Face09lmms-lab-eval /MMVP MMVP (Multimodal Visual Patterns) Benchmark This is a corrected version of the MMVP benchmark, re-hosted by lmms-lab-eval for use with lmms-eval. Why this copy? The original MMVP/MMVP dataset was uploaded in imagefolder format, which only exposes the image column. The text annotations (Question, Options, Correct Answer, Index) from the accompanying Questions.csv were not loaded into the dataset, making it unusable for evaluation. This version reconstructs the complete… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-eval/MMVP.imagevisual-question-answeringn<1K0 likes2.4k downloads7mo agoHugging Face10lmms-eval /ActivityNetQAtext1K<n<10K7 likes2.3k downloads2y agoHugging Face11lmms-eval /VideoMMMUgatedThis dataset contains the data for the paper Video-MMMU: Evaluating Knowledge Acquisition from Multi-Discipline Professional Videos. Video-MMMU is a multi-modal, multi-disciplinary benchmark designed to assess LMMs' ability to acquire and utilize knowledge from videos. Project page: https://videommmu.github.io/ Leaderboard (last updated: 07 Feb, 2025) Model Overall Perception Comprehension Adaptation Δknowledge Human Expert 74.44 84.33 78.67 60.33 +33.1… See the full description on the dataset page: https://huggingface.co/datasets/lmms-eval/VideoMMMU.imagen<1K17 likes2.1k downloads1y agoHugging Face12yifanzhang114 /MME-RealWorld-Lmms-evaltext10K<n<100K1 likes1.9k downloads2y agoHugging Face13oscarqjh /MindCube_lmmseval MindCube LMMs Eval Dataset This dataset is formatted for use with lmms-eval framework. Dataset Schema Column Type Description id string Unique identifier for each sample (format: {split}_{scene_id}_{question_id}) category list[string] Category labels (e.g., ["perpendicular", "P-O", "meanwhile", "self"]) type string Question type (e.g., "1_frame", "2_frame", "3_frame", "general") meta_info list[list[string]] Metadata about scene objects and their spatial… See the full description on the dataset page: https://huggingface.co/datasets/oscarqjh/MindCube_lmmseval.image10K<n<100K1 likes1.6k downloads8mo agoHugging Face14CaraJ /MathVerse-lmmseval Dataset Card for MathVerse This is the version for lmms-eval. This shares the same data with the official dataset. Dataset Description Paper Information Dataset Examples Leaderboard Citation Dataset Description The capabilities of Multi-modal Large Language Models (MLLMs) in visual math problem-solving remain insufficiently evaluated and understood. We investigate current benchmarks to incorporate excessive visual content within textual questions, which potentially… See the full description on the dataset page: https://huggingface.co/datasets/CaraJ/MathVerse-lmmseval.imagemultiple-choice1K<n<10K2 likes1.5k downloads2y agoHugging Face15lmms-eval /PerceptionTest_Valaudio10K<n<100K1 likes1.2k downloads2y agoHugging Face16lmms-eval /charades_statext1K<n<10K6 likes1.1k downloads2y agoHugging Face17lmms-eval /VideoChatGPTtext1K<n<10K13 likes862 downloads2y agoHugging Face18lmms-eval /video-tt Towards Video Thinking Test (Video-TT): A Holistic Benchmark for Advanced Video Reasoning and Understanding Video-TT comprises 1,000 YouTube videos, each paired with one open-ended question and four adversarial questions designed to probe visual and narrative complexity. Paper: https://arxiv.org/abs/2507.15028 Project page: https://zhangyuanhan-ai.github.io/video-tt/ 🚀 What's New [2025.03] We release the benchmark! 1. Why Do We Need a New… See the full description on the dataset page: https://huggingface.co/datasets/lmms-eval/video-tt.text10K<n<100K5 likes854 downloads1y agoHugging Face19lmms-lab-eval /Spatial457image10K<n<100K0 likes820 downloads8mo agoHugging Face20lmms-lab-eval /egotempo EgoTempo Full-set metadata for lmms-eval task egotempo. Annotation source: https://raw.githubusercontent.com/google-research-datasets/egotempo/main/egotempo_openQA.json Raw video location: Ego4D clips (license-gated), clip id in clip_id. Expected local media root for evaluation: $EGOTEMPO_VIDEO_DIR or $HF_HOME/egotempo. textn<1K0 likes820 downloads7mo agoHugging Face21lmms-eval /TOMATOtabular1K<n<10K0 likes802 downloads1y agoHugging Face22oscarqjh /ViewSpatial_lmmsevalimage1K<n<10K1 likes781 downloads9mo agoHugging Face23HuggingFaceM4 /lmms-eval-embeddingsThese are the precomputed embeddings of 66 image-text benchmarks from the lmms-eval framework, intended for use in the large-scale-image-deduplication repository as mentioned in the FineVision Blogpost They can be downloaded using the cli: hf download HuggingFaceM4/lmms-eval-embeddings --local-dir embeddings --repo-type dataset feature-extraction2 likes722 downloads1y agoHugging Face24oscarqjh /3DSRBench_lmmseval 3DSRBench (lmms-eval compatible) This is a reformatted version of 3DSRBench for compatibility with lmms-eval. Dataset Description 3DSRBench is a comprehensive 3D spatial reasoning benchmark that evaluates the 3D spatial reasoning capabilities of Large Multimodal Models (LMMs). It includes 2,100 VQAs on MS-COCO images and 672 on multi-view synthetic images rendered from HSSD. Subsets This dataset contains two subsets: 1. 3dsr… See the full description on the dataset page: https://huggingface.co/datasets/oscarqjh/3DSRBench_lmmseval.imagevisual-question-answering10K<n<100K0 likes721 downloads8mo agoHugging Face25yifanzhang114 /MME-RealWorld-lite-lmms-eval 2024.11.14 🌟 MME-RealWorld now has a lite version (50 samples per task, or all if fewer than 50) for inference acceleration, which is also supported by VLMEvalKit and Lmms-eval. 2024.09.03 🌟 MME-RealWorld is now supported in the VLMEvalKit and Lmms-eval repository, enabling one-click evaluation—give it a try!" 2024.08.20 🌟 We are very proud to launch MME-RealWorld, which contains 13K high-quality images, annotated by 32 volunteers, resulting in 29K question-answer pairs that cover 43… See the full description on the dataset page: https://huggingface.co/datasets/yifanzhang114/MME-RealWorld-lite-lmms-eval.text1K<n<10K1 likes607 downloads2y agoHugging Face26yifanzhang114 /MME-RealWorld-CN-Lmms-evaltext1K<n<10K1 likes547 downloads2y agoHugging Face27lmms-eval /PerceptionTestaudio10K<n<100K2 likes460 downloads2y agoHugging Face28lmms-eval /MMVUtext1K<n<10K2 likes446 downloads1y agoHugging Face29lmms-lab-eval /HLE-Verified HLE-Verified (HF-native) This dataset is a Hugging Face-native conversion of skylenage/HLE-Verified at revision becad9f339dfce27df0ebb38e55dabef12ca5735. Why this exists The source dataset stores nested verification fields with mixed runtime types (for example 0/1/"uncertain"), which breaks strict Arrow JSON parsing in datasets.load_dataset. This converted dataset normalizes those fields and publishes split-ready JSONL files for direct use in lmms_eval. Split… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-eval/HLE-Verified.text1K<n<10K0 likes443 downloads7mo agoHugging Face30lmms-eval /VATEXtext1K<n<10K1 likes383 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.