CoolFace
Datasetpublic

SPEAK-PP/sinhala-spelling-correction-already-corrected-pairs

Sinhala ASR Prediction-Reference Dataset (3000 no-numbers) Dataset Description This dataset contains sentence pairs for spelling correction: dyslexic_sentence: noisy / predicted text clean_sentence: clean reference text Dataset Statistics Split Samples Train 2,400 Eval 300 Test 300 Total 3,000 Usage from datasets import load_dataset dataset =… See the full description on the dataset page: https://huggingface.co/datasets/SPEAK-PP/sinhala-spelling-correction-already-corrected-pairs.

sourceHugging Facemitupdated 7mo agoView on Hugging Face
0likes10downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
SPEAK-PP/sinhala-spelling-correction-already-corrected-pairs · CoolFace