datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
rvl-cdip-document-classification
rvl-cdip-document-classification
This dataset is created from original aharley/rvl_cdip dataset using this notebook
Dataset Summary
This dataset consists of 8992 grayscale images in 16 classes, with 562 images per class.
There are 8000 training images(500 image per class) and 992 test images(62 images per class).
The images are sized so their largest dimension does not exceed 1000 pixels.
document-classification-benchmark
Document Classification Benchmark (open-vocab, zero-shot)
Given a document image and an arbitrary set of text labels, which one is right? A held-out, zero-shot,
open-vocabulary evaluation for document-type classification — labels are supplied at inference, not baked
into a head. Test split only; not for training. Every image is drawn from a permissively-licensed,
redistributable source.
Powers the
document-classification-leaderboard
and evaluates document-classification-v2… See the full description on the dataset page: https://huggingface.co/datasets/nutrientdocs/document-classification-benchmark.
