CoolFace
Datasetpublic

tum-nlp/cannot-dataset

Compilation of ANnotated, Negation-Oriented Text-pairs Dataset Card for CANNOT Dataset Summary CANNOT is a dataset that focuses on negated textual pairs. It currently contains 77,376 samples, of which roughly of them are negated pairs of sentences, and the other half are not (they are paraphrased versions of each other). The most frequent negation that appears in the dataset is verbal negation (e.g., will → won't), although it also contains pairs… See the full description on the dataset page: https://huggingface.co/datasets/tum-nlp/cannot-dataset.

sourceHugging Facecc-by-sa-4.0updated 11mo agoView on Hugging Face
0likes46downloads
11 commits on main
b56975811mo ago

update citation

MiriUll
bb311cd3y ago

Update README.md

MiriUll
dad57db3y ago

Added citation

MiriUll
27be3ab3y ago

Add further dataset stats

dmlls
59c284b3y ago

Add size_categories to dataset metadata

dmlls
ea702333y ago

Add license to dataset metadata

dmlls
f1f54663y ago

Update README.md

Diego Miguel
7497b3d3y ago

Update dataset filename

Diego Miguel
900ffee3y ago

Upload negation_dataset_v1.0.tsv

dmlls
99366563y ago

Create README.md

dmlls
75553a03y ago

initial commit

dmlls