CoolFace
Datasetpublic

biglam/loc_beyond_words

Dataset Card for Beyond Words Dataset Summary The Beyond Words dataset is a crowdsourced collection of bounding box annotations on World War I-era historical newspaper pages from the Library of Congress’s Chronicling America collection. Volunteers marked seven types of visual content — photographs, illustrations, maps, comics, editorial cartoons, headlines, and advertisements — enabling the training of the visual content recognition model behind the Newspaper… See the full description on the dataset page: https://huggingface.co/datasets/biglam/loc_beyond_words.

sourceHugging Facecc0-1.0updated 1y agoView on Hugging Face
15likes112downloads
11 commits on main
6c7f5fb1y ago

switch to parquet version of dataset (#3)

davanstrien
f3d04b61y ago

Upload dataset (#2)

davanstrien
9d8779c1y ago

Update README.md (#1)

davanstrien
e6adb4b4y ago

Update README.md

davanstrien
888fe054y ago

Update loc_beyond_words.py

davanstrien
3a030454y ago

Update README.md

davanstrien
67c5a1d4y ago

Update README.md

davanstrien
7c7133d4y ago

Update README.md

davanstrien
d9f9db54y ago

add readme

davanstrien
f241e214y ago

draft dataset

davanstrien
8172af94y ago

initial commit

davanstrien