lopentu/Chinese-Wordnet-SemCor
Chinese Wordnet SemCor Dataset Summary This dataset is designed for the task of Word Sense Disambiguation (WSD) for common Chinese words, specifically focusing on words identified as "difficult" (having more than 10 senses) within Chinese Wordnet (CWN) 2.0. It originates from the annotation dataset described in Section 3.1 of the paper "Resolving Regular Polysemy in Named Entities." The original dataset consisted of 28,836 example sentences where a target… See the full description on the dataset page: https://huggingface.co/datasets/lopentu/Chinese-Wordnet-SemCor.
141
Update README.md
update tags
update tags
add tags
add licensing information
Update README.md
Upload dataset
add citation
initial commit
