moaminsharifi/fa-wiki-spell-checker
Persian / Farsi Wikipedia Corpus for Spell Checking Tasks Overview The Wikipedia Corpus is an open source dataset specifically designed for use in spell checking tasks. It is available on huggingface and can be accessed and utilized by anyone interested in improving spell checking algorithms. Formula chance of being % normal sentences >=2% manipulation <=98% each time with random function we create a new random number for each line… See the full description on the dataset page: https://huggingface.co/datasets/moaminsharifi/fa-wiki-spell-checker.
This repository belongs to moaminsharifi on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
