opendoc
Datasets
All datasets matching “opendoc”OpenDocVQA-Corpus-ExtendedOpenDoc-Pdf-Preview
OpenDoc-Pdf-Preview
OpenDoc-Pdf-Preview is a compact visual preview dataset containing 6,000 high-resolution document images extracted from PDFs. This dataset is designed for Image-to-Text tasks such as document OCR pretraining, layout understanding, and multimodal document analysis.
Dataset Summary
Modality: Image-to-Text
Content Type: PDF-based document previews
Number of Samples: 6,000
Language: English
Format: Parquet
Split: train only
Size: 606 MB
License: Apache… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/OpenDoc-Pdf-Preview.OpenDocVQA
Dataset Card for OpenDocVQA
This is a training and evaluation QA data file for VDocRAG, a new RAG framework that can directly understand diverse real-world documents purely from visual features.
Dataset Description
OpenDocVQA is the first unified collection of open-domain document visual question answering datasets, encompassing diverse document types and formats.
Supported Tasks and Leaderboards
Given a large collection of document images and a question, the… See the full description on the dataset page: https://huggingface.co/datasets/NTT-hil-insight/OpenDocVQA.OpenDocVQA-Corpus
Dataset Card for OpenDocVQA
This is a training and evaluation corpus data file for VDocRAG, a new RAG framework that can directly understand diverse real-world documents purely from visual features.
Dataset Description
OpenDocVQA is the first unified collection of open-domain document visual question answering datasets, encompassing diverse document types and formats.
Supported Tasks and Leaderboards
Given a large collection of document images and a question… See the full description on the dataset page: https://huggingface.co/datasets/NTT-hil-insight/OpenDocVQA-Corpus.OpenDocVQA-ExtendedOpendoc2-Analysis-Recognition
Opendoc2-Analysis-Recognition Dataset
Overview
The Opendoc2-Analysis-Recognition dataset is a collection of data designed for tasks involving image analysis and recognition. It is suitable for various machine learning tasks, including image-to-text conversion, text classification, and image feature extraction.
Dataset Details
Modalities: Likely includes images and associated labels (specific modalities can be confirmed on the dataset's page).
Languages:… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Opendoc2-Analysis-Recognition.
