CCB/cis5300-text-classification
Complex Word Identification (CIS 5300) Dataset Description This dataset supports the Complex Word Identification (CWI) task: given a word in context, predict whether it is complex (likely to be difficult for non-native speakers, children, or people with reading disabilities) or simple. CWI is the first step in lexical simplification — the task of rewriting text to make it more accessible. Before you can simplify a word, you need to identify which words need… See the full description on the dataset page: https://huggingface.co/datasets/CCB/cis5300-text-classification.
Upload README.md with huggingface_hub
Upload README.md with huggingface_hub
Upload folder using huggingface_hub
Delete data/ngram_counts.txt.gz with huggingface_hub
Delete data/cwi2018_news_test.tsv with huggingface_hub
Delete data/complex_words_training.txt with huggingface_hub
Delete data/complex_words_test_unlabeled.txt with huggingface_hub
Delete data/complex_words_test_mini.txt with huggingface_hub
Delete data/complex_words_development_mini_unlabeled.txt with huggingface_hub
Delete data/complex_words_development_mini.txt with huggingface_hub
Delete data/complex_words_development.txt with huggingface_hub
Delete data/complex_biomed_test.tsv with huggingface_hub
Upload folder using huggingface_hub
initial commit
