CoolFace
6 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01xandery /OpenDocVQA-Corpus-Extendedimage10K<n<100K0 likes307 downloads9mo agoHugging Face02prithivMLmods /OpenDoc-Pdf-Preview OpenDoc-Pdf-Preview OpenDoc-Pdf-Preview is a compact visual preview dataset containing 6,000 high-resolution document images extracted from PDFs. This dataset is designed for Image-to-Text tasks such as document OCR pretraining, layout understanding, and multimodal document analysis. Dataset Summary Modality: Image-to-Text Content Type: PDF-based document previews Number of Samples: 6,000 Language: English Format: Parquet Split: train only Size: 606 MB License: Apache… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/OpenDoc-Pdf-Preview.documentimage-to-text1K<n<10K2 likes283 downloads1y agoHugging Face03NTT-hil-insight /OpenDocVQA-Corpusgated Dataset Card for OpenDocVQA This is a training and evaluation corpus data file for VDocRAG, a new RAG framework that can directly understand diverse real-world documents purely from visual features. Dataset Description OpenDocVQA is the first unified collection of open-domain document visual question answering datasets, encompassing diverse document types and formats. Supported Tasks and Leaderboards Given a large collection of document images and a question… See the full description on the dataset page: https://huggingface.co/datasets/NTT-hil-insight/OpenDocVQA-Corpus.imagevisual-question-answering100K<n<1M4 likes73 downloads1y agoHugging Face04prithivMLmods /Opendoc2-Analysis-Recognition Opendoc2-Analysis-Recognition Dataset Overview The Opendoc2-Analysis-Recognition dataset is a collection of data designed for tasks involving image analysis and recognition. It is suitable for various machine learning tasks, including image-to-text conversion, text classification, and image feature extraction. Dataset Details Modalities: Likely includes images and associated labels (specific modalities can be confirmed on the dataset's page). Languages:… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Opendoc2-Analysis-Recognition.imageimage-to-text1K<n<10K2 likes49 downloads2y agoHugging Face05prithivMLmods /OpenDoc-Null-6K OpenDoc-Null-6K The OpenDoc-Null-6K dataset is curated for tasks related to image-to-text recognition, particularly for scanned document images and OCR (Optical Character Recognition) use cases. It contains over 6,900 images in a structured imagefolder format suitable for training models on document parsing, PDF image understanding, and layout/text extraction tasks. Attribute Value Task Image-to-Text Modality Image Format ImageFolder Language English License… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/OpenDoc-Null-6K.imageimage-to-text1K<n<10K2 likes28 downloads1y agoHugging Face06prithivMLmods /Opendoc1-Analysis-Recognition Opendoc1-Analysis-Recognition Dataset Overview The Opendoc1-Analysis-Recognition dataset is designed for tasks involving image-to-text, text classification, and image feature extraction. It contains images paired with class labels, making it suitable for vision-language tasks. Dataset Details Modalities: Image Languages: English Size: Approximately 1,000 samples (n=1K) Tags: image, analysis, vision-language License: Apache 2.0 Tasks This dataset… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Opendoc1-Analysis-Recognition.imageimage-to-textn<1K3 likes27 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.