tum-nlp/cannot-dataset
Compilation of ANnotated, Negation-Oriented Text-pairs Dataset Card for CANNOT Dataset Summary CANNOT is a dataset that focuses on negated textual pairs. It currently contains 77,376 samples, of which roughly of them are negated pairs of sentences, and the other half are not (they are paraphrased versions of each other). The most frequent negation that appears in the dataset is verbal negation (e.g., will → won't), although it also contains pairs… See the full description on the dataset page: https://huggingface.co/datasets/tum-nlp/cannot-dataset.
update citation
Update README.md
Added citation
Add further dataset stats
Add size_categories to dataset metadata
Add license to dataset metadata
Update README.md
Update dataset filename
Upload negation_dataset_v1.0.tsv
Create README.md
initial commit
