CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01lmms-lab-encoder /DocVQA Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of DocVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @article{mathew2020docvqa, title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)}… See the full description on the dataset page: https://huggingface.co/datasets/lmms-lab-encoder/DocVQA.image10K<n<100K88 likes38k downloads2y agoHugging Face02hf-internal-testing /fixtures_docvqaThis dataset includes 2 document images of the DocVQA dataset. They are used for testing the LayoutLMv2FeatureExtractor + LayoutLMv2Processor inside the HuggingFace Transformers library. More specifically, they are used in tests/test_feature_extraction_layoutlmv2.py and tests/test_processor_layoutlmv2.py. imagen<1K0 likes4.7k downloads1y agoHugging Face03pixparse /docvqa-single-page-questions Dataset Card for DocVQA Dataset Dataset Summary DocVQA dataset is a document dataset introduced in Mathew et al. (2021) consisting of 50,000 questions defined on 12,000+ document images. Please visit the challenge page (https://rrc.cvc.uab.es/?ch=17) and paper (https://arxiv.org/abs/2007.00398) for further information. Usage This dataset can be used with current releases of Hugging Face datasets library. Here is an example using a custom collator to bundle… See the full description on the dataset page: https://huggingface.co/datasets/pixparse/docvqa-single-page-questions.imagequestion-answering10K<n<100K11 likes2.9k downloads2y agoHugging Face04lmms-lab-encoder /MP-DocVQAimage10K<n<100K8 likes2.3k downloads3y agoHugging Face05vidore /docvqa_test_subsampled_beirBEIR version of vidore/docvqa_test_subsampled. imagedocument-question-answering1K<n<10K0 likes1.9k downloads1y agoHugging Face06openbmb /VisRAG-Ret-Test-MP-DocVQA Dataset Description This is a VQA dataset based on Industrial Documents from MP-DocVQA dataset from MP-DocVQA. Load the dataset from datasets import load_dataset import csv def load_beir_qrels(qrels_file): qrels = {} with open(qrels_file) as f: tsvreader = csv.DictReader(f, delimiter="\t") for row in tsvreader: qid = row["query-id"] pid = row["corpus-id"] rel = int(row["score"]) if qid in qrels:… See the full description on the dataset page: https://huggingface.co/datasets/openbmb/VisRAG-Ret-Test-MP-DocVQA.image1K<n<10K1 likes1.7k downloads2y agoHugging Face07VLR-CVC /DocVQA-2026 DocVQA 2026 | ICDAR2026 Competition on Multimodal Reasoning over Documents in Multiple Domains Building upon previous DocVQA benchmarks, this evaluation dataset introduces challenging reasoning questions over a diverse collection of documents spanning eight domains, including business reports, scientific papers, slides, posters, maps, comics, infographics, and engineering drawings. By expanding coverage to new document domains and… See the full description on the dataset page: https://huggingface.co/datasets/VLR-CVC/DocVQA-2026.imagevisual-question-answeringn<1K74 likes1.4k downloads17d agoHugging Face08mteb /docvqa_test_subsampled_beirBEIR version of vidore/docvqa_test_subsampled. imagedocument-question-answering1K<n<10K0 likes1.1k downloads8mo agoHugging Face09vidore /docvqa_test_subsampled Dataset Description This is the test set taken from the DocVQA dataset. It includes collected images from the UCSF Industry Documents Library. Questions and answers were manually annotated. Example of data (see viewer) Data Curation To ensure homogeneity across our benchmarked datasets, we subsampled the original test set to 500 pairs and renamed the different columns. Load the dataset from datasets import load_dataset ds =… See the full description on the dataset page: https://huggingface.co/datasets/vidore/docvqa_test_subsampled.imagedocument-question-answeringn<1K6 likes908 downloads1y agoHugging Face10nielsr /docvqa_1200_examplesimage1K<n<10K12 likes769 downloads4y agoHugging Face11cmarkea /doc-vqa Dataset description The doc-vqa Dataset integrates images from the Infographic_vqa dataset sourced from HuggingFaceM4 The Cauldron dataset, as well as images from the dataset AFTDB (Arxiv Figure Table Database) curated by cmarkea. This dataset consists of pairs of images and corresponding text, with each image linked to an average of five questions and answers available in both English and French. These questions and answers were generated using Gemini 1.5 Pro, thereby… See the full description on the dataset page: https://huggingface.co/datasets/cmarkea/doc-vqa.imagevisual-question-answering10K<n<100K19 likes615 downloads2y agoHugging Face12AbdulMuqtadir /DocVQA_Processed_Datasetimage10K<n<100K0 likes567 downloads3y agoHugging Face13jdchandana /DocVQA Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of DocVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @article{mathew2020docvqa, title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)}… See the full description on the dataset page: https://huggingface.co/datasets/jdchandana/DocVQA.image10K<n<100K0 likes542 downloads4mo agoHugging Face14jnamlee /DocVQA Large-scale Multi-modality Models Evaluation Suite Accelerating the development of large-scale multi-modality models (LMMs) with lmms-eval 🏠 Homepage | 📚 Documentation | 🤗 Huggingface Datasets This Dataset This is a formatted version of DocVQA. It is used in our lmms-eval pipeline to allow for one-click evaluations of large multi-modality models. @article{mathew2020docvqa, title={DocVQA: A Dataset for VQA on Document Images. CoRR abs/2007.00398 (2020)}… See the full description on the dataset page: https://huggingface.co/datasets/jnamlee/DocVQA.image10K<n<100K0 likes477 downloads1mo agoHugging Face15vidore /docvqa_trainimagedocument-question-answering10K<n<100K1 likes443 downloads1y agoHugging Face16danjacobellis /docvqaimage10K<n<100K0 likes327 downloads2y agoHugging Face17vikhyatk /docvqaimage10K<n<100K1 likes316 downloads2y agoHugging Face18albertklorer /DocVQAimagequestion-answering10K<n<100K0 likes313 downloads8mo agoHugging Face19jinaai /docvqa Creation This dataset is build upon the corresponding dataset from the ViDoRe Benchmark. For more information regarding the filtering please read our paper or this discussion on github. Disclaimer This dataset may contain publicly available images or text data. All data is provided for research and educational purposes only. If you are the rights holder of any content and have concerns regarding intellectual property or copyright, please contact us at "support-data… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/docvqa.imagen<1K0 likes281 downloads1y agoHugging Face20jinaai /docvqa_beirThis is a copy of https://huggingface.co/datasets/jinaai/docvqa reformatted into the BEIR format. For any further information like license, please refer to the original dataset. Disclaimer This dataset may contain publicly available images or text data. All data is provided for research and educational purposes only. If you are the rights holder of any content and have concerns regarding intellectual property or copyright, please contact us at "support-data (at) jina.ai" for… See the full description on the dataset page: https://huggingface.co/datasets/jinaai/docvqa_beir.imagen<1K0 likes280 downloads1y agoHugging Face21emrekuruu /MP-DocVQA MP-DocVQA MP-DocVQA is one of the 11 retrieval benchmarks used in RetrievalRouter: Joint Modality and Architecture Selection for Document Retrieval (EMNLP 2026). Each record pairs a rendered page image, a query, and the page's extracted text, supporting both text-based and multimodal retrieval evaluation. 📄 Paper: https://arxiv.org/pdf/2608.25625 💻 Code: https://github.com/emrekuruu/retrieval-router 🤗 Collection: https://huggingface.co/collections/emrekuruu/retrieval-router… See the full description on the dataset page: https://huggingface.co/datasets/emrekuruu/MP-DocVQA.imagetext-retrieval1K<n<10K0 likes239 downloads28d agoHugging Face22prashanthpillai /docvqa_train_and_val Dataset Card for "docvqa_train_and_val" More Information needed tabular10K<n<100K2 likes229 downloads3y agoHugging Face23RIPS-Goog-23 /DocVQAtabular10K<n<100K0 likes223 downloads3y agoHugging Face24Sharka /DocVQA_layoutLM Dataset Card for "DocVQA_layoutLM" More Information needed tabular10K<n<100K0 likes206 downloads3y agoHugging Face25Ssunbell /boostcamp-docvqa-v4 Dataset Card for "boostcamp-docvqa-v4" More Information needed tabular10K<n<100K0 likes171 downloads4y agoHugging Face26mm-eval /DocVQAimage10K<n<100K0 likes169 downloads2mo agoHugging Face27Ssunbell /boostcamp-docvqa Dataset Card for "boostcamp-docvqa" More Information needed tabular10K<n<100K0 likes167 downloads4y agoHugging Face28Ssunbell /boostcamp-docvqa-v5 Dataset Card for "boostcamp-docvqa-v5" More Information needed tabular10K<n<100K1 likes156 downloads4y agoHugging Face29nielsr /docvqa_1200_examples_donutimage1K<n<10K10 likes143 downloads4y agoHugging Face30Ssunbell /boostcamp-docvqa-marker Dataset Card for "boostcamp-docvqa-marker" More Information needed tabular10K<n<100K0 likes135 downloads4y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.