CoolFace
Datasetpublic

Weaxs/csc

Dataset for CSC 中文纠错数据集 Dataset Description Chinese Spelling Correction (CSC) is a task to detect and correct misspelled characters in Chinese texts. 共计 120w 条数据,以下是数据来源 数据集 语料 链接 SIGHAN+Wang271K 拼写纠错数据集 SIGHAN+Wang271K(27万条) https://huggingface.co/datasets/shibing624/CSC ECSpell 拼写纠错数据集 包含法律、医疗、金融等领域 https://github.com/Aopolin-Lv/ECSpell CGED 语法纠错数据集 仅包含了2016和2021年的数据集… See the full description on the dataset page: https://huggingface.co/datasets/Weaxs/csc.

sourceHugging Faceapache-2.0updated 3y agoView on Hugging Face
7likes125downloads
23 commits on main
033274f3y ago

update

Weaxs
236a5ee3y ago

upload backup

Weaxs
93039703y ago

Delete grammar

Weaxs
c316f813y ago

Delete SIGHAN+Wang271K

Weaxs
e038e223y ago

Delete NLPCC

Weaxs
24a61313y ago

Delete ECSpell

Weaxs
0f43a5d3y ago

Delete CGED

Weaxs
4f1d1b83y ago

Delete NLG

Weaxs
2538fbe3y ago

Update .gitattributes

Weaxs
9469e603y ago

Update README.md

Weaxs
de8f5513y ago

Update README.md

Weaxs
59708bc3y ago

Upload 3 files

Weaxs
fa04c533y ago

Rename validate.jsonl to validation.jsonl

Weaxs
d07d1be3y ago

Upload validate.jsonl

Weaxs
f3f58353y ago

Update README.md

Weaxs
37cf2543y ago

Update README.md

Weaxs
87349fd3y ago

Update README.md

Weaxs
4546e343y ago

gpt merged dataset

Weaxs
f7c0e6a3y ago

del nlpcc18

Weaxs
d938d9f3y ago

upload dataset

Weaxs
c5513843y ago

Upload 3 files

Weaxs
4bea3de3y ago

Upload 4 files

Weaxs
ad956f13y ago

initial commit

Weaxs