CoolFace
Datasetpublic

agbalu/KabLiterary

KabLiterary A 28,907-paragraph, 657,713-word corpus of classical world literature in canonical Kabyle Latin orthography, from the AƔBALU project. Eleven works — translated or originally composed in Kabyle — spanning epic poetry, gothic fiction, philosophical prose, folktales, military strategy, and 19th-century novels. The corpus is designed for language modelling, vocabulary probing, fine-tuning, and literary translation benchmarking where long-form, high-register Kabyle text… See the full description on the dataset page: https://huggingface.co/datasets/agbalu/KabLiterary.

sourceHugging Facecc-by-sa-4.0updated 3d agoView on Hugging Face
0likes38downloads
filedev.parquet44 KBdownload
filetest.parquet106 KBdownload
filetrain.parquet2.8 MBdownload

agbalu/KabLiterary · main · files are served by the source, never re-hosted here