CoolFace
10 results

opendoc

xandery /OpenDocVQA-Corpus-Extendedimage10K<n<100K0 likes307 downloads9mo agoHugging FaceprithivMLmods /OpenDoc-Pdf-Preview OpenDoc-Pdf-Preview OpenDoc-Pdf-Preview is a compact visual preview dataset containing 6,000 high-resolution document images extracted from PDFs. This dataset is designed for Image-to-Text tasks such as document OCR pretraining, layout understanding, and multimodal document analysis. Dataset Summary Modality: Image-to-Text Content Type: PDF-based document previews Number of Samples: 6,000 Language: English Format: Parquet Split: train only Size: 606 MB License: Apache… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/OpenDoc-Pdf-Preview.documentimage-to-text1K<n<10K2 likes283 downloads1y agoHugging FaceNTT-hil-insight /OpenDocVQA Dataset Card for OpenDocVQA This is a training and evaluation QA data file for VDocRAG, a new RAG framework that can directly understand diverse real-world documents purely from visual features. Dataset Description OpenDocVQA is the first unified collection of open-domain document visual question answering datasets, encompassing diverse document types and formats. Supported Tasks and Leaderboards Given a large collection of document images and a question, the… See the full description on the dataset page: https://huggingface.co/datasets/NTT-hil-insight/OpenDocVQA.textvisual-question-answering10K<n<100K3 likes207 downloads1y agoHugging FaceNTT-hil-insight /OpenDocVQA-Corpusgated Dataset Card for OpenDocVQA This is a training and evaluation corpus data file for VDocRAG, a new RAG framework that can directly understand diverse real-world documents purely from visual features. Dataset Description OpenDocVQA is the first unified collection of open-domain document visual question answering datasets, encompassing diverse document types and formats. Supported Tasks and Leaderboards Given a large collection of document images and a question… See the full description on the dataset page: https://huggingface.co/datasets/NTT-hil-insight/OpenDocVQA-Corpus.imagevisual-question-answering100K<n<1M4 likes73 downloads1y agoHugging Facexandery /OpenDocVQA-Extendedtext10K<n<100K0 likes62 downloads1mo agoHugging FaceprithivMLmods /Opendoc2-Analysis-Recognition Opendoc2-Analysis-Recognition Dataset Overview The Opendoc2-Analysis-Recognition dataset is a collection of data designed for tasks involving image analysis and recognition. It is suitable for various machine learning tasks, including image-to-text conversion, text classification, and image feature extraction. Dataset Details Modalities: Likely includes images and associated labels (specific modalities can be confirmed on the dataset's page). Languages:… See the full description on the dataset page: https://huggingface.co/datasets/prithivMLmods/Opendoc2-Analysis-Recognition.imageimage-to-text1K<n<10K2 likes49 downloads2y agoHugging Face