CoolFace
20 results

sarc

iabufarha /ar_sarcasm Dataset Card for ArSarcasm Dataset Summary ArSarcasm is a new Arabic sarcasm detection dataset. The dataset was created using previously available Arabic sentiment analysis datasets (SemEval 2017 and ASTD) and adds sarcasm and dialect labels to them. The dataset contains 10,547 tweets, 1,682 (16%) of which are sarcastic. For more details, please check the paper From Arabic Sentiment Analysis to Sarcasm Detection: The ArSarcasm Dataset Supported Tasks and… See the full description on the dataset page: https://huggingface.co/datasets/iabufarha/ar_sarcasm.texttext-classification10K<n<100K18 likes490 downloads3y agoHugging FaceMedOtter /Soft-tissue-Sarcoma Soft-tissue-Sarcoma (STS) A TCIA collection of 51 patients with histologically proven soft-tissue sarcoma of the extremities, each imaged with joint pre-treatment FDG-PET/CT and MRI and contoured by an expert radiation oncologist. Collected at McGill University Health Centre (Montreal) and published with Vallières et al., Phys Med Biol 2015. The original study built a radiomics model predicting lung metastases from joint PET/MRI texture features; 19 of the 51 patients developed… See the full description on the dataset page: https://huggingface.co/datasets/MedOtter/Soft-tissue-Sarcoma.imageimage-segmentationn<1K0 likes374 downloads2mo agoHugging Faceraquiba /Sarcasm_News_HeadlinePast studies in Sarcasm Detection mostly make use of Twitter datasets collected using hashtag based supervision but such datasets are noisy in terms of labels and language. Furthermore, many tweets are replies to other tweets and detecting sarcasm in these requires the availability of contextual tweets. To overcome the limitations related to noise in Twitter datasets, this Headlines dataset for Sarcasm Detection is collected from two news website. TheOnion aims at producing sarcastic versions… See the full description on the dataset page: https://huggingface.co/datasets/raquiba/Sarcasm_News_Headline.text10K<n<100K6 likes274 downloads4y agoHugging Facemarcbishara /sarcasm-on-redditCopied from: Sarcasm on Reddit. https://www.kaggle.com/datasets/danofer/sarcasm Which in turn came from: @unpublished{SARC, authors={Mikhail Khodak and Nikunj Saunshi and Kiran Vodrahalli}, title={A Large Self-Annotated Corpus for Sarcasm}, url={https://arxiv.org/abs/1704.05579}, year=2017 } license: mit language: - en tabular1M<n<10M1 likes255 downloads10mo agoHugging Facegabrielaltay /tcga-sarc-tabular-open TCGA-SARC — Tabular (Open Access) Open-access TCGA-SARC data from the NCI Genomic Data Commons, reshaped into one table per GDC data_type. Clinical, biospecimen and every open molecular modality for this cohort, in one place, queryable without downloading a single .tar or parsing a single TSV. GDC data release: Data Release 46.0 - August 10, 2026 Built: 2026-09-12 04:16:39 UTC Scope: one TCGA project — see [the family][repo] for the others from datasets import load_dataset… See the full description on the dataset page: https://huggingface.co/datasets/gabrielaltay/tcga-sarc-tabular-open.tabular100M<n<1B0 likes227 downloads15d agoHugging Facetasksource /figlang2020-sarcasmhttps://github.com/EducationalTestingService/sarcasm @inproceedings{ghosh-etal-2020-report, title = "A Report on the 2020 Sarcasm Detection Shared Task", author = "Ghosh, Debanjan and Vajpayee, Avijit and Muresan, Smaranda", booktitle = "Proceedings of the Second Workshop on Figurative Language Processing", month = jul, year = "2020", address = "Online", publisher = "Association for Computational Linguistics", url =… See the full description on the dataset page: https://huggingface.co/datasets/tasksource/figlang2020-sarcasm.3 likes213 downloads3y agoHugging Face