CoolFace
Datasetpublic

agbalu/KabLiterary

KabLiterary A 28,907-paragraph, 657,713-word corpus of classical world literature in canonical Kabyle Latin orthography, from the AƔBALU project. Eleven works — translated or originally composed in Kabyle — spanning epic poetry, gothic fiction, philosophical prose, folktales, military strategy, and 19th-century novels. The corpus is designed for language modelling, vocabulary probing, fine-tuning, and literary translation benchmarking where long-form, high-register Kabyle text… See the full description on the dataset page: https://huggingface.co/datasets/agbalu/KabLiterary.

sourceHugging Facecc-by-sa-4.0updated 3d agoView on Hugging Face
0likes38downloads
2 commits on main
b842e333d ago

Upload folder using huggingface_hub

ainouche-abderahmane
8c2b6743d ago

initial commit

ainouche-abderahmane