CoolFace
Datasetpublic

shibing624/CSC

Dataset Card for CSC 中文拼写纠错数据集 Repository: https://github.com/shibing624/pycorrector Dataset Description Chinese Spelling Correction (CSC) is a task to detect and correct misspelled characters in Chinese texts. CSC is challenging since many Chinese characters are visually or phonologically similar but with quite different semantic meanings. 中文拼写纠错数据集,共27万条,是通过原始SIGHAN13、14、15年数据集和Wang271k数据集合并整理后得到,json格式,带错误字符位置信息。 Original Dataset Summary… See the full description on the dataset page: https://huggingface.co/datasets/shibing624/CSC.

sourceHugging Faceapache-2.0updated 3y agoView on Hugging Face
37likes214downloads

Nothing at this path on main. The folder may be empty, or the revision may not exist.

shibing624/CSC · main · files are served by the source, never re-hosted here

shibing624/CSC · CoolFace