CoolFace
Datasetpublic

abdelhaqueidali/Kabyle-Latin-to-Tifinagh-Parallel-Corpus

Dataset Card for Kabyle Latin-to-Tifinagh Parallel Corpus This dataset provides a parallel corpus of the Kabyle language (Taqbaylit), pairing native Latin-based orthography with automated, context-aware Amazigh script transliterations. It is built by processing raw text data through a rule-based algorithmic pipeline designed to enforce strict orthographic purity, manage contextual phonetic mutations, and isolate foreign vocabulary. Dataset Details… See the full description on the dataset page: https://huggingface.co/datasets/abdelhaqueidali/Kabyle-Latin-to-Tifinagh-Parallel-Corpus.

sourceHugging Facecc-by-2.0updated 3mo agoView on Hugging Face
0likes52downloads

abdelhaqueidali/Kabyle-Latin-to-Tifinagh-Parallel-Corpus · main · files are served by the source, never re-hosted here