CoolFace
20 results

Sarcasm

iabufarha /ar_sarcasm Dataset Card for ArSarcasm Dataset Summary ArSarcasm is a new Arabic sarcasm detection dataset. The dataset was created using previously available Arabic sentiment analysis datasets (SemEval 2017 and ASTD) and adds sarcasm and dialect labels to them. The dataset contains 10,547 tweets, 1,682 (16%) of which are sarcastic. For more details, please check the paper From Arabic Sentiment Analysis to Sarcasm Detection: The ArSarcasm Dataset Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/iabufarha/ar_sarcasm.texttext-classification10K<n<100K18 likes505 downloads3y agoHugging Faceraquiba /Sarcasm_News_HeadlinePast studies in Sarcasm Detection mostly make use of Twitter datasets collected using hashtag based supervision but such datasets are noisy in terms of labels and language. Furthermore, many tweets are replies to other tweets and detecting sarcasm in these requires the availability of contextual tweets. To overcome the limitations related to noise in Twitter datasets, this Headlines dataset for Sarcasm Detection is collected from two news website. TheOnion aims at producing sarcastic versions… See the full description on the dataset page: https://huggingface.co/datasets/raquiba/Sarcasm_News_Headline.text10K<n<100K6 likes260 downloads4y agoHugging FaceSpellOnYou /kor_sarcasm Dataset Card for Korean Sarcasm Detection Dataset Summary The Korean Sarcasm Dataset was created to detect sarcasm in text, which can significantly alter the original meaning of a sentence. 9319 tweets were collected from Twitter and labeled for sarcasm or not_sarcasm. These tweets were gathered by querying for: 역설, 아무말, 운수좋은날, 笑, 뭐래 아닙니다, 그럴리없다, 어그로, irony sarcastic, and sarcasm. The dataset was pre-processed by removing the keyword hashtag, urls and mentions of the user… See the full description on the dataset page: https://huggingface.co/datasets/SpellOnYou/kor_sarcasm.texttext-classification1K<n<10K5 likes169 downloads2y agoHugging Facetasksource /figlang2020-sarcasmhttps://github.com/EducationalTestingService/sarcasm @inproceedings{ghosh-etal-2020-report, title = "A Report on the 2020 Sarcasm Detection Shared Task", author = "Ghosh, Debanjan and Vajpayee, Avijit and Muresan, Smaranda", booktitle = "Proceedings of the Second Workshop on Figurative Language Processing", month = jul, year = "2020", address = "Online", publisher = "Association for Computational Linguistics", url =… See the full description on the dataset page: https://huggingface.co/datasets/tasksource/figlang2020-sarcasm.3 likes161 downloads3y agoHugging FaceArrebol-yzq /Metaphor_Driven_Sarcasm_dataset_source Metaphor_Driven_Sarcasm_dataset_source Dataset Description This is a Chinese metaphor-driven sarcasm detection dataset, containing 235,909 texts with a four‑layer annotation structure. The dataset is converted from the BSD‑MM‑V3 dataset and is specifically designed for studying the role of metaphor in sarcastic expressions. Dataset Summary Volume: 235,909 texts Language: Chinese Annotation Levels: 4 (sarcasm category, sarcasm style, metaphor… See the full description on the dataset page: https://huggingface.co/datasets/Arrebol-yzq/Metaphor_Driven_Sarcasm_dataset_source.text-classification100K<n<1M1 likes144 downloads21d agoHugging FaceCreativeLang /SARC_Sarcasm SARC_Sarcasm Dataset Summary A large corpus for sarcasm research and for training and evaluating systems for sarcasm detection is presented. The corpus comprises 1.3 million sarcastic statements, a quantity that is tenfold more substantial than any preceding dataset, and includes many more instances of non-sarcastic statements. This allows for learning in both balanced and unbalanced label regimes. Each statement is self-annotated; that is to say, sarcasm is labeled by… See the full description on the dataset page: https://huggingface.co/datasets/CreativeLang/SARC_Sarcasm.tabular10M<n<100M3 likes95 downloads3y agoHugging Face