CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01yale-nlp /MMVU MMVU: Measuring Expert-Level Multi-Discipline Video Understanding 🌐 Homepage • 🥇 Leaderboard • 📖 Paper • 🤗 Data 📰 News 2025-01-21: We are excited to release the MMVU paper, dataset, and evaluation code! 👋 Overview Why MMVU Benchmark? Despite the rapid progress of foundation models in both text-based and image-based expert reasoning, there is a clear gap in evaluating these models’ capabilities in specialized-domain video understanding.… See the full description on the dataset page: https://huggingface.co/datasets/yale-nlp/MMVU.textvideo-text-to-text1K<n<10K59 likes3.6k downloads2y agoHugging Face02lmms-lab-eval /MMVP MMVP (Multimodal Visual Patterns) Benchmark This is a corrected version of the MMVP benchmark, re-hosted by lmms-lab-eval for use with lmms-eval. Why this copy? The original MMVP/MMVP dataset was uploaded in imagefolder format, which only exposes the image column. The text annotations (Question, Options, Correct Answer, Index) from the accompanying Questions.csv were not loaded into the dataset, making it unusable for evaluation. This version reconstructs the complete… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-eval/MMVP.imagevisual-question-answeringn<1K0 likes2.4k downloads7mo agoHugging Face03lmms-lab-encoder /MMVet Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of MM-Vet. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @misc{yu2023mmvet, title={MM-Vet: Evaluating Large Multimodal Models for Integrated Capabilities}… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-encoder/MMVet.imagen<1K5 likes1.6k downloads3y agoHugging Face04whyu /mm-vetPaper: https://arxiv.org/abs/2308.02490 imagen<1K3 likes539 downloads2y agoHugging Face05lmms-eval /MMVUtext1K<n<10K2 likes446 downloads1y agoHugging Face06nimapourjafar /mm_vqav2image10K<n<100K0 likes357 downloads2y agoHugging Face07mm-vl /mcot_r1_mcq_66kimage10K<n<100K0 likes197 downloads1y agoHugging Face08nimapourjafar /mm_visual7wimage10K<n<100K0 likes191 downloads2y agoHugging Face09whyu /mm-vet-v2Paper: https://arxiv.org/abs/2408.00765 imagen<1K0 likes146 downloads2y agoHugging Face10mm-eval /MMVetimagen<1K0 likes106 downloads2mo agoHugging Face11mm-vl /mcot_r1_vqa_66kimage10K<n<100K0 likes96 downloads1y agoHugging Face12czty /MMV-dataset MMV-dataset Reproducible research splits and full-training subsets derived from long-video benchmarks. No video, audio, or subtitle media is included. Hosted data Component Train Validation Grouping unit License LongVideoBench labeled validation set 401 936 video_id CC BY-NC-SA 4.0 EgoTempo open-ended QA 150 350 original Ego4D video UID CC BY 4.0 Unified supervised records 1,084 2,529 inherited mixed; see per-row license The LongVideoBench… See the full description on the dataset page: https://huggingface.co/datasets/czty/MMV-dataset.tabularquestion-answering1K<n<10K0 likes94 downloads22d agoHugging Face13mm-vl /x2x_rft_22kimage10K<n<100K1 likes79 downloads1y agoHugging Face14mm-vl /x2x_rft_16kimage10K<n<100K1 likes62 downloads1y agoHugging Face15sming256 /MMVU_MCtextn<1K0 likes61 downloads11mo agoHugging Face16zhouyik /MMVMBench_VQAtext1K<n<10K0 likes43 downloads1y agoHugging Face17mm-eval /MMVPimagen<1K0 likes43 downloads2mo agoHugging Face18mteb /MMVU-VQA MMVUVideoCentricQA An MTEB dataset Massive Text Embedding Benchmark MMVU is an expert-level, multi-discipline video understanding benchmark with questions spanning 27 subjects across Science, Healthcare, Humanities & Social Sciences, and Engineering. Each multiple-choice example pairs a specialized-domain video with a question and 5 candidate answers. The task is formulated as multiple-choice retrieval: given the (video, question) pair, retrieve the correct candidate. Used the public… See the full description on the dataset page: https://huggingface.co/datasets/mteb/MMVU-VQA.textvisual-question-answering1K<n<10K0 likes42 downloads2mo agoHugging Face19nimapourjafar /mm_vistextimage1K<n<10K2 likes38 downloads2y agoHugging Face20nimapourjafar /mm_visualmrcimage1K<n<10K0 likes37 downloads2y agoHugging Face21lccshunli /MMVUtextn<1K0 likes37 downloads6mo agoHugging Face22HuggingFaceM4 /MM_VET_modif Dataset Card for "MM_VET_modif" MM-VET Benchmark imagen<1K1 likes34 downloads3y agoHugging Face23lhpku20010120 /MM-Verify-Datatext10K<n<100K3 likes34 downloads2y agoHugging Face24nimapourjafar /mm_vqaradimagen<1K0 likes28 downloads2y agoHugging Face25mm-eval /MM-Vet-v2 ⚠️ DEPRECATED This repository is deprecated — use mm-eval/MMVet-v2 instead. mm-eval/MM-Vet-v2 and mm-eval/MMVet-v2 are duplicate uploads of the same benchmark (MM-Vet v2, 517 identical rows — same ids, questions, and answers; confirmed in the 2026-07-07 org audit). Per the owner's decision the newer conversion MMVet-v2 is the canonical copy. The data here is kept unchanged for reproducibility of past runs; do not use it for new evaluations. imagen<1K0 likes27 downloads2mo agoHugging Face26Wissam42 /MMVU-VQA MMVUVideoCentricQA An MTEB dataset Massive Text Embedding Benchmark MMVU is an expert-level, multi-discipline video understanding benchmark with questions spanning 27 subjects across Science, Healthcare, Humanities & Social Sciences, and Engineering. Each multiple-choice example pairs a specialized-domain video with a question and 5 candidate answers. The task is formulated as multiple-choice retrieval: given the (video, question) pair, retrieve the correct candidate. Used the public… See the full description on the dataset page: https://huggingface.co/datasets/Wissam42/MMVU-VQA.textvisual-question-answering1K<n<10K0 likes27 downloads2mo agoHugging Face27mm-eval /MMVet-v2imagen<1K0 likes25 downloads2mo agoHugging Face28FoteiniTag /MMVetimagen<1K0 likes25 downloads2mo agoHugging Face29MaoSong2022 /MMVP MMVP Benchmark refactor MMVP to support VLMEalKit Benchmark Information number of questions: 300 question type: multiple choice question question format: image + text Reference VLMEvalKit MMVP textmultiple-choicen<1K0 likes21 downloads1y agoHugging Face30pkulium /MMVetimagen<1K0 likes17 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.