CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01GonzaloA /fake_news TODO: Add YAML tags here. Copy-paste the tags obtained with the online tagging app: https://huggingface.co/spaces/huggingface/datasets-tagging annotations_creators: - no-annotation language_creators: - found language: - en license: - unknown multilinguality: - monolingual size_categories: - 30k<n<50k source_datasets: - original task_categories: - text-classification task_ids: - fact-checking - intent-classification pretty_name: GonzaloA / Fake News Dataset Card for… See the full description on the dataset page: https://huggingface.co/datasets/GonzaloA/fake_news.tabular10K<n<100K28 likes2.2k downloads4y agoHugging Face02genesisqu /fake-real-newstext10K<n<100K0 likes601 downloads4y agoHugging Face03mrm8488 /fake-newstext10K<n<100K1 likes347 downloads5y agoHugging Face04community-datasets /fake_news_english Dataset Card for Fake News English Dataset Summary This dataset contains URLs of news articles classified as either fake or satire. The articles classified as fake also have the URL of a rebutting article. Supported Tasks and Leaderboards [More Information Needed] Languages English Dataset Structure Data Instances { "article_number": 102 , "url_of_article":… See the full description on the dataset page: https://huggingface.co/datasets/community-datasets/fake_news_english.texttext-classificationn<1K3 likes217 downloads2y agoHugging Face05mariagrandury /fake_news_corpus_spanish Fake News Corpus Spanish Citation Gómez-Adorno, H., Posadas-Durán, J. P., Enguix, G. B., & Capetillo, C. P. (2021). Overview of FakeDeS at IberLEF 2021: Fake News Detection in Spanish Shared Task. Procesamiento del Lenguaje Natural, 67, 223-231. Aragón, M. E., Jarquín, H., Gómez, M. M. Y., Escalante, H. J., Villaseñor-Pineda, L., Gómez-Adorno, H., ... & Posadas-Durán, J. P. (2020, September). Overview of mex-a3t at iberlef 2020: Fake news and aggressiveness analysis in… See the full description on the dataset page: https://huggingface.co/datasets/mariagrandury/fake_news_corpus_spanish.texttext-classificationn<1K2 likes213 downloads2y agoHugging Face06ErfanMoosaviMonazzah /fake-news-detection-dataset-EnglishThis is a cleaned and splitted version of this dataset (https://www.kaggle.com/datasets/sadikaljarif/fake-news-detection-dataset-english) Labels: Fake News: 0 Real News: 1 You can find the cleansing script at: https://github.com/ErfanMoosaviMonazzah/Fake-News-Detection tabulartext-classification10K<n<100K5 likes209 downloads4y agoHugging Face07Ambrosio1994 /real-and-fake-newstext10K<n<100K1 likes178 downloads1y agoHugging Face08Trinisha /fake_or_real_news Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/Trinisha/fake_or_real_news.text1K<n<10K0 likes164 downloads3y agoHugging Face09Annabelleabbott /real-fake-news-workshoptextn<1K3 likes162 downloads5y agoHugging Face10nanyy1025 /covid_fake_newsConstraint@AAAI2021 - COVID19 Fake News Detection in English @misc{patwa2020fighting, title={Fighting an Infodemic: COVID-19 Fake News Dataset}, author={Parth Patwa and Shivam Sharma and Srinivas PYKL and Vineeth Guptha and Gitanjali Kumari and Md Shad Akhtar and Asif Ekbal and Amitava Das and Tanmoy Chakraborty}, year={2020}, eprint={2011.03327}, archivePrefix={arXiv}, primaryClass={cs.CL} } texttext-classification10K<n<100K2 likes160 downloads4y agoHugging Face11community-datasets /urdu_fake_news Dataset Card for Bend the Truth (Urdu Fake News) Dataset Summary [More Information Needed] Supported Tasks and Leaderboards [More Information Needed] Languages [More Information Needed] Dataset Structure Data Instances [More Information Needed] Data Fields news: a string in urdu label: the label indicating whethere the provided news is real or fake. category: The intent of the news being presented. The available 5… See the full description on the dataset page: https://huggingface.co/datasets/community-datasets/urdu_fake_news.texttext-classificationn<1K3 likes159 downloads2y agoHugging Face12Edds /spanish-fake-news-fixed Spanish Fake News Fixed Este dataset contiene noticias etiquetadas en español, reparado para corregir saltos de línea internos. text10K<n<100K0 likes158 downloads3mo agoHugging Face13sayalaruano /FakeNewsSpanish_Kaggle2This dataset was obtained from: https://www.kaggle.com/datasets/zulanac/fake-and-real-news textn<1K1 likes151 downloads5y agoHugging Face14shahafvl /john_fake_newstabular100K<n<1M0 likes120 downloads2y agoHugging Face15vikasgautam2003 /Fake_and_Real_newstext1K<n<10K0 likes92 downloads1y agoHugging Face16BeardedJohn /FakeNewstext10K<n<100K2 likes91 downloads4y agoHugging Face17mohammadjavadpirhadi /fake-news-detection-dataset-english Dataset Card for "fake-news-detection-dataset-english" More Information needed texttext-classification10K<n<100K0 likes88 downloads4y agoHugging Face18suwaimyo /fakenews-fil-classification Fakenews_fil_Classification Deduplicated copy of kornwtp/fakenews-fil-classification. Splits split rows train 3,005 text1K<n<10K0 likes86 downloads29d agoHugging Face19LittleFish-Coder /Fake_News_GossipCopDataset Source: Ahren09/MMSoc_GossipCop This is a copied and reformatted version of the Ahren09/MMSoc_GossipCop text: text of the article (str) bert_embeddings: (768, ) roberta_embeddings: (768, ) label: (int) 0: real 1: fake Datasets Distribution: Train: 9988 (real: 7955, fake: 2033) Test: 2672 (real: 2169, 503) tabular10K<n<100K0 likes81 downloads2mo agoHugging Face20noahgift /fake-newstext1K<n<10K0 likes80 downloads4y agoHugging Face21winterForestStump /fake-news-detector-euvsdisinfodata from https://euvsdisinfo.eu/ text1K<n<10K0 likes80 downloads2y agoHugging Face22shahafvl /kobby_fake_newsDataset: ikekobby/40-percent-cleaned-preprocessed-fake-real-news tabulartext-classification100K<n<1M0 likes79 downloads2y agoHugging Face23ikekobby /40-percent-cleaned-preprocessed-fake-real-newsKaggle based dataset for text classification task. The data has been cleaned and processed for preparation into any model for classification based tasks. This is just 40% of the entire dataset. text10K<n<100K1 likes73 downloads4y agoHugging Face24andyP /fake_news_en_opensources Dataset Card for "Fake News Opensources" Dataset Description Homepage: https://github.com/AndyTheFactory/FakeNewsDataset Repository: https://github.com/AndyTheFactory/FakeNewsDataset Point of Contact: Andrei Paraschiv Dataset Summary a consolidated and cleaned up version of the opensources Fake News dataset Fake News Corpus comprises 8,529,090 individual articles, classified into 12 classes: reliable, unreliable, political, bias, fake, conspiracy… See the full description on the dataset page: https://huggingface.co/datasets/andyP/fake_news_en_opensources.texttext-classification1M<n<10M2 likes68 downloads3y agoHugging Face25LittleFish-Coder /Fake_News_KDD2020Dataset Source: Fake News Detection Challenge KDD 2020 This is a copied and reformatted version of the Fake News Detection Challenge KDD 2020. We use the raw train.csv from the official Kaggle Dataset and split the data into train and test sets. text: text of the article (str) embeddings: BERT embeddings (768, ) label: (int) 1: fake 0: true Datasets Distribution: Train: 4487 Test: 499 texttext-classification1K<n<10K0 likes66 downloads2y agoHugging Face26AyushiAyushi3017 /fake-news-detector-datasettext10K<n<100K1 likes63 downloads10mo agoHugging Face27Cartinoe5930 /Politifact_fake_newstabular10K<n<100K2 likes61 downloads3y agoHugging Face28fake-news-UFG /central_de_fatos Central de Fatos Dataset Summary In recent times, the interest for research dissecting the dissemination and prevention of misinformation in the online environment has spiked dramatically. Given that scenario, a recurring obstacle is the unavailability of public datasets containing fact-checked instances. In this work, we performed an extensive data collection of such instances from the better part of all major internationally recognized Brazilian fact-checking agencies.… See the full description on the dataset page: https://huggingface.co/datasets/fake-news-UFG/central_de_fatos.texttext-classification10K<n<100K1 likes61 downloads3y agoHugging Face29pushpdeep /fake_news_combinedLabel Description 0 : Fake, 1 : Real tabular10K<n<100K0 likes53 downloads3y agoHugging Face30toni5rovic /bcms-fake-news-articlestexttext-classification10K<n<100K0 likes52 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.