CoolFace
Datasetpublic

SPEAK-PP/sinhala-spelling-correction-already-corrected-pairs

Sinhala ASR Prediction-Reference Dataset (3000 no-numbers) Dataset Description This dataset contains sentence pairs for spelling correction: dyslexic_sentence: noisy / predicted text clean_sentence: clean reference text Dataset Statistics Split Samples Train 2,400 Eval 300 Test 300 Total 3,000 Usage from datasets import load_dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/SPEAK-PP/sinhala-spelling-correction-already-corrected-pairs.

sourceHugging Facemitupdated 7mo agoView on Hugging Face
0likes10downloads
3 commits on main
dd7f99a7mo ago

Add dataset card

kcdw
5f02af67mo ago

Initial upload: train/eval/test dataset

kcdw
81d52157mo ago

initial commit

kcdw