CoolFace
Datasetpublic

CreativeLang/wps_chinese_simile

WPS - Chinese Simile Dataset Summary Chinese Simile (CS) Dataset This dataset is constructed and based on the online free-access fictions that are tagged with sci-fi, urban novel, love story, youth, etc. All similes are extracted by rich regular expression, and the extraction precision is estimated as 92% by labelling 500 random extracted samples. Further data filtering as well as processing is truly encouraged! The data split in paper is as follows (You could… See the full description on the dataset page: https://huggingface.co/datasets/CreativeLang/wps_chinese_simile.

sourceHugging Facecc-by-2.0updated 3y agoView on Hugging Face
0likes25downloads
4 commits on main
4f315f43y ago

Update README.md

liyucheng
dd581be3y ago

Update README.md

liyucheng
ce92a803y ago

Upload 3 files

liyucheng
f65569d3y ago

initial commit

liyucheng