CoolFace
Datasetpublic

moaminsharifi/fa-wiki-spell-checker

Persian / Farsi Wikipedia Corpus for Spell Checking Tasks Overview The Wikipedia Corpus is an open source dataset specifically designed for use in spell checking tasks. It is available on huggingface and can be accessed and utilized by anyone interested in improving spell checking algorithms. Formula chance of being % normal sentences >=2% manipulation <=98% each time with random function we create a new random number for each line… See the full description on the dataset page: https://huggingface.co/datasets/moaminsharifi/fa-wiki-spell-checker.

sourceHugging Facepddlupdated 3y agoView on Hugging Face
6likes21downloads
9 commits on main
3c2c7dd3y ago

Update README.md

moaminsharifi
3b4761b3y ago

Update README.md

moaminsharifi
f05d5d13y ago

Upload fawiki-spell-checker.csv

moaminsharifi
b48aa233y ago

Delete wiki-fa-articles-spell-checker.txt

moaminsharifi
98005fa3y ago

Update README.md

moaminsharifi
765e8ad3y ago

Upload wiki-fa-articles-spell-checker.txt

moaminsharifi
e7902d63y ago

Update README.md

moaminsharifi
c843cc03y ago

Update README.md

moaminsharifi
16108e03y ago

initial commit

moaminsharifi