CoolFace
Datasetpublicgated

djelia/bm-text-normalization

bm-text-normalization Bambara (Bamanankan) orthographic normalisation: map a non-standard spelling to its standard form. 4,877 short phrase-level pairs in a single config, bamadaba. Load from datasets import load_dataset train = load_dataset("djelia/bm-text-normalization", "bamadaba", split="train") dev = load_dataset("djelia/bm-text-normalization", "bamadaba", split="dev") test = load_dataset("djelia/bm-text-normalization", "bamadaba", split="test") # rows… See the full description on the dataset page: https://huggingface.co/datasets/djelia/bm-text-normalization.

sourceHugging Faceupdated 2mo agoView on Hugging Face
0likes13downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.