CoolFace
Datasetpublic

cis-lmu/GlotStoryBook

Dataset Description Story Books for 180 ISO-639-3 codes. The Parallel ID or parallel_id can be used to find the parallel documents in different languages and build a parallel dataset. This dataset consists of 2 subsets: default, which consists of 4 publishers: asp: African Storybook pb: Pratham Books lcb: Little Cree Books lida: LIDA Stories nalibali, which comes from Nal'ibali stories. Usage (HF Loader) default: from datasets import load_dataset dataset… See the full description on the dataset page: https://huggingface.co/datasets/cis-lmu/GlotStoryBook.

sourceHugging Faceccupdated 5d agoView on Hugging Face
9likes193downloads
settings

This repository belongs to cis-lmu on Hugging Face.

CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.

nameGlotStoryBook
visibilitypublic
licencecc
gatedno
ownercis-lmu
Account settings
cis-lmu/GlotStoryBook · CoolFace