cis-lmu/GlotStoryBook
Dataset Description Story Books for 180 ISO-639-3 codes. The Parallel ID or parallel_id can be used to find the parallel documents in different languages and build a parallel dataset. This dataset consists of 2 subsets: default, which consists of 4 publishers: asp: African Storybook pb: Pratham Books lcb: Little Cree Books lida: LIDA Stories nalibali, which comes from Nal'ibali stories. Usage (HF Loader) default: from datasets import load_dataset dataset… See the full description on the dataset page: https://huggingface.co/datasets/cis-lmu/GlotStoryBook.
update yaml
add nalibali readme.
add nalibali
add nalibali
remove adx, khg.
remove khg.
remove adx.
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
add alternative scripts
Upload GlotStoryBook.csv
index: false
add parallel id
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Update README.md
Upload GlotStoryBook.csv
initial commit
