CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01TIGER-Lab /MMEB-eval Massive Multimodal Embedding Benchmark We compile a large set of evaluation tasks to understand the capabilities of multimodal embedding models. This benchmark covers 4 meta tasks and 36 datasets meticulously selected for evaluation. The dataset is published in our paper VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks. Dataset Usage For each dataset, we have 1000 examples for evaluation. Each example contains a query and a set of… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/MMEB-eval.image10K<n<100K16 likes3.7k downloads2y agoHugging Face02TIGER-Lab /MMEB-train Massive Multimodal Embedding Benchmark The training data split used for training VLM2Vec models in the paper VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks (ICLR 2025). MMEB benchmark covers 4 meta tasks and 36 datasets meticulously selected for evaluating capabilities of multimodal embedding models. During training, we utilize 20 out of the 36 datasets. For evaluation, we assess performance on the 20 in-domain (IND) datasets and the remaining 16… See the full description on the dataset page: https://huggingface.co/datasets/TIGER-Lab/MMEB-train.image1M<n<10M18 likes3.5k downloads2y agoHugging Face03VLM2Vec /MMEB-V3 MMEB-V3: Measuring the Performance Gaps of Omni-Modality Embedding Models 🌐 Website | GitHub | 🏆 Leaderboard | 📖 MMEB-V3 Paper | 📖 MMEB-V2 Paper | 📖 MMEB-V1 Paper | 🤗 Models Introduction MMEB-V3 is a comprehensive benchmark for evaluating omni-modality embedding models across text, image, video, audio, visual-document, and agent-centric retrieval scenarios. Building upon MMEB-V1 and MMEB-V2, MMEB-V3 adds 111 new tasks, resulting in 190 evaluation tasks in… See the full description on the dataset page: https://huggingface.co/datasets/VLM2Vec/MMEB-V3.textfeature-extraction5 likes2.6k downloads2mo agoHugging Face04tianyiqi1 /MMEB-train-DocVQA-images MMEB-train — DocVQA / Train images Backup copy of the DocVQA Train image split used in MMEB-train (VLM2Vec) training. Original folder structure is preserved: files live under DocVQA/Train/. Files: 78,926 JPG images Size: ~15 GB image1K<n<10K0 likes933 downloads2mo agoHugging Face05MrZilinXiao /MMEB_train_with_imageimage1M<n<10M0 likes860 downloads1y agoHugging Face06LucasLima /MME-Benchmark-pt Avaliação - MME-Perception Estrutura do Diretório main ├── MME_Benchmark │ ├── artwork │ │ ├── images │ │ │ ├── 1.jpg │ │ │ ├── 2.jpg │ │ │ ├── ... │ │ ├── question_answers_YN │ │ │ ├── 1.txt │ │ │ ├── 2.txt │ │ │ ├── ... │ ├── celebrity │ ├── code_reasoning │ ├── ... ├── calculation.py ├── translate_MME Estrutura dos Arquivos TXT Cada arquivo num.txt contém as perguntas correspondentes à imagem num.jpg.… See the full description on the dataset page: https://huggingface.co/datasets/LucasLima/MME-Benchmark-pt.imagen<1K0 likes589 downloads2y agoHugging Face07laughatwill /TIGER-Lab_MMEB-train MMEB Training Dataset (Lance Format) This is a Lance-format version of the TIGER-Lab/MMEB-train dataset, optimized for efficient storage and fast random access. The original dataset is used for training VLM2Vec models in the paper VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks (ICLR 2025). Directory Structure TIGER-Lab_MMEB-train/ └── data/ ├── A-OKVQA/ │ ├── train.lance │ ├── original.lance │ └── diverse.lance… See the full description on the dataset page: https://huggingface.co/datasets/laughatwill/TIGER-Lab_MMEB-train.image1M<n<10M0 likes214 downloads8mo agoHugging Face08MrZilinXiao /MMEB-eval-ChartQA-beir-v3image1K<n<10K0 likes44 downloads1y agoHugging Face09MrZilinXiao /MMEB-eval-ChartQA-beirimage1K<n<10K0 likes43 downloads1y agoHugging Face10MrZilinXiao /MMEB-eval-DocVQA-beirimage1K<n<10K0 likes42 downloads1y agoHugging Face11MrZilinXiao /MMEB-eval-Wiki-SS-NQ-beir-v3image10K<n<100K0 likes37 downloads1y agoHugging Face12MrZilinXiao /MMEB-eval-VisDial-beirimage1K<n<10K0 likes35 downloads1y agoHugging Face13MrZilinXiao /MMEB-eval-OVEN-beir-v3image1K<n<10K0 likes33 downloads1y agoHugging Face14MrZilinXiao /MMEB-eval-A-OKVQA-beirimage1K<n<10K0 likes32 downloads1y agoHugging Face15MrZilinXiao /MMEB-eval-MSCOCO_i2t-beirimage1K<n<10K0 likes32 downloads1y agoHugging Face16MrZilinXiao /MMEB-eval-ObjectNet-beirimage1K<n<10K0 likes31 downloads1y agoHugging Face17MrZilinXiao /MMEB-eval-GQA-beir-v2image1K<n<10K0 likes31 downloads1y agoHugging Face18MrZilinXiao /MMEB-eval-OK-VQA-beirimage1K<n<10K0 likes30 downloads1y agoHugging Face19MrZilinXiao /MMEB-eval-ImageNet-R-beir-v3image1K<n<10K0 likes30 downloads1y agoHugging Face20MrZilinXiao /MMEB-eval-FashionIQ-beir-v3image1K<n<10K0 likes30 downloads1y agoHugging Face21MrZilinXiao /MMEB-eval-RefCOCO-Matching-beir-v2image1K<n<10K0 likes29 downloads1y agoHugging Face22MrZilinXiao /MMEB-eval-GQA-beir-v3image1K<n<10K0 likes28 downloads1y agoHugging Face23ManukyanD /MMEB-train-subsampledimage100K<n<1M0 likes26 downloads2y agoHugging Face24MrZilinXiao /MMEB-eval-ImageNet-A-beirimage1K<n<10K0 likes26 downloads1y agoHugging Face25MrZilinXiao /MMEB-eval-TextVQA-beir-v3image1K<n<10K0 likes26 downloads1y agoHugging Face26MrZilinXiao /MMEB-eval-WebQA-beirimage1K<n<10K0 likes24 downloads1y agoHugging Face27MrZilinXiao /MMEB-eval-ScienceQA-beirimage1K<n<10K0 likes24 downloads1y agoHugging Face28MrZilinXiao /MMEB-eval-TextVQA-beirimage1K<n<10K0 likes23 downloads1y agoHugging Face29MrZilinXiao /MMEB-eval-WebQA-beir-v2image1K<n<10K0 likes23 downloads1y agoHugging Face30MrZilinXiao /MMEB-eval-VisDial-beir-v3image1K<n<10K0 likes23 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.