CoolFace
Datasetpublic

moaminsharifi/fa-wiki-spell-checker

Persian / Farsi Wikipedia Corpus for Spell Checking Tasks Overview The Wikipedia Corpus is an open source dataset specifically designed for use in spell checking tasks. It is available on huggingface and can be accessed and utilized by anyone interested in improving spell checking algorithms. Formula chance of being % normal sentences >=2% manipulation <=98% each time with random function we create a new random number for each line… See the full description on the dataset page: https://huggingface.co/datasets/moaminsharifi/fa-wiki-spell-checker.

sourceHugging Facepddlupdated 3y agoView on Hugging Face
6likes21downloads

moaminsharifi/fa-wiki-spell-checker · main · files are served by the source, never re-hosted here