CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01victor /real-or-fake-fake-jobposting-predictiontabular10K<n<100K5 likes4.8k downloads4y agoHugging Face02gplsi /fake_job_postings_balanced_en 🧠 BALANCED_FAKE_JOB_POSTINGS_EN Dataset 📘 Overview This dataset is a balanced English version of the original Fake Job Postings dataset from Kaggle: Real or Fake? Fake Job Posting Prediction. It contains 1,730 job postings, equally divided between fraudulent (fake) and non-fraudulent (real) listings. All text fields remain in English, preserving the semantic meaning and structure of the original dataset. Only balancing was performed — no translation or additional… See the full description on the dataset page: https://huggingface.co/datasets/gplsi/fake_job_postings_balanced_en.tabulartext-classification1K<n<10K0 likes1.4k downloads9mo agoHugging Face03genesisqu /fake-real-newstext10K<n<100K0 likes604 downloads4y agoHugging Face04serenityyyyy /fake_job_postings_balanced_en 🧠 BALANCED_FAKE_JOB_POSTINGS_EN Dataset 📘 Overview This dataset is a balanced English version of the original Fake Job Postings dataset from Kaggle: Real or Fake? Fake Job Posting Prediction. It contains 1,730 job postings, equally divided between fraudulent (fake) and non-fraudulent (real) listings. All text fields remain in English, preserving the semantic meaning and structure of the original dataset. Only balancing was performed — no translation or additional… See the full description on the dataset page: https://huggingface.co/datasets/serenityyyyy/fake_job_postings_balanced_en.tabulartext-classification1K<n<10K0 likes507 downloads6mo agoHugging Face05haja17 /real-or-fake-fake-jobposting-predictiontabular10K<n<100K0 likes480 downloads7mo agoHugging Face06mrm8488 /fake-newstext10K<n<100K1 likes344 downloads5y agoHugging Face07nehalvatss /real-or-fake-fake-jobposting-predictiontabular10K<n<100K0 likes278 downloads3mo agoHugging Face08Aurthor /fake_job_post_predictiontabular10K<n<100K2 likes233 downloads2y agoHugging Face09mariagrandury /fake_news_corpus_spanish Fake News Corpus Spanish Citation Gómez-Adorno, H., Posadas-Durán, J. P., Enguix, G. B., & Capetillo, C. P. (2021). Overview of FakeDeS at IberLEF 2021: Fake News Detection in Spanish Shared Task. Procesamiento del Lenguaje Natural, 67, 223-231. Aragón, M. E., Jarquín, H., Gómez, M. M. Y., Escalante, H. J., Villaseñor-Pineda, L., Gómez-Adorno, H., ... & Posadas-Durán, J. P. (2020, September). Overview of mex-a3t at iberlef 2020: Fake news and aggressiveness analysis in… See the full description on the dataset page: https://huggingface.co/datasets/mariagrandury/fake_news_corpus_spanish.texttext-classificationn<1K2 likes213 downloads2y agoHugging Face10ErfanMoosaviMonazzah /fake-news-detection-dataset-EnglishThis is a cleaned and splitted version of this dataset (https://www.kaggle.com/datasets/sadikaljarif/fake-news-detection-dataset-english) Labels: Fake News: 0 Real News: 1 You can find the cleansing script at: https://github.com/ErfanMoosaviMonazzah/Fake-News-Detection tabulartext-classification10K<n<100K5 likes212 downloads4y agoHugging Face11ADILSHARMA2007 /real-or-fake-fake-jobposting-predictiontabular10K<n<100K0 likes189 downloads28d agoHugging Face12recogna-nlp /FakeRecogna FakeRecogna FakeRecogna is a dataset comprised of real and fake news. The real news is not directly linked to fake news and vice-versa, which could lead to a biased classification. The news collection was performed by crawlers developed for mining pages of well-known and of great national importance agency news. The web crawlers were developed based on each analyzed webpage, where the extracted information is first separated into categories and then grouped by dates. The plurality… See the full description on the dataset page: https://huggingface.co/datasets/recogna-nlp/FakeRecogna.texttext-classification10K<n<100K3 likes177 downloads3y agoHugging Face13Trinisha /fake_or_real_news Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/Trinisha/fake_or_real_news.text1K<n<10K0 likes164 downloads3y agoHugging Face14nanyy1025 /covid_fake_newsConstraint@AAAI2021 - COVID19 Fake News Detection in English @misc{patwa2020fighting, title={Fighting an Infodemic: COVID-19 Fake News Dataset}, author={Parth Patwa and Shivam Sharma and Srinivas PYKL and Vineeth Guptha and Gitanjali Kumari and Md Shad Akhtar and Asif Ekbal and Amitava Das and Tanmoy Chakraborty}, year={2020}, eprint={2011.03327}, archivePrefix={arXiv}, primaryClass={cs.CL} } texttext-classification10K<n<100K2 likes161 downloads4y agoHugging Face15recogna-nlp /fakerecogna2-abstrativa FakeRecogna 2.0 - Abstractive FakeRecogna 2.0 presents the extension for the FakeRecogna dataset in the context of fake news detection. FakeRecogna includes real and fake news texts collected from online media and ten fact-checking sources in Brazil. An important aspect is the lack of relation between the real and fake news samples, i.e., they are not mutually related to each other to avoid intrinsic bias in the data. The Dataset The fake news collection was performed on… See the full description on the dataset page: https://huggingface.co/datasets/recogna-nlp/fakerecogna2-abstrativa.tabulartext-classification10K<n<100K2 likes156 downloads1y agoHugging Face16Edds /spanish-fake-news-fixed Spanish Fake News Fixed Este dataset contiene noticias etiquetadas en español, reparado para corregir saltos de línea internos. text10K<n<100K0 likes149 downloads3mo agoHugging Face17sayalaruano /FakeNewsSpanish_Kaggle2This dataset was obtained from: https://www.kaggle.com/datasets/zulanac/fake-and-real-news textn<1K1 likes145 downloads5y agoHugging Face18recogna-nlp /fakerecogna2-extrativa FakeRecogna 2.0 Extractive FakeRecogna 2.0 presents the extension for the FakeRecogna dataset in the context of fake news detection. FakeRecogna includes real and fake news texts collected from online media and ten fact-checking sources in Brazil. An important aspect is the lack of relation between the real and fake news samples, i.e., they are not mutually related to each other to avoid intrinsic bias in the data. The Dataset The fake news collection was performed on… See the full description on the dataset page: https://huggingface.co/datasets/recogna-nlp/fakerecogna2-extrativa.tabulartext-classification10K<n<100K1 likes135 downloads1y agoHugging Face19BeardedJohn /FakeNewstext10K<n<100K2 likes92 downloads4y agoHugging Face20vikasgautam2003 /Fake_and_Real_newstext1K<n<10K0 likes91 downloads1y agoHugging Face21winterForestStump /fake-news-detector-euvsdisinfodata from https://euvsdisinfo.eu/ text1K<n<10K0 likes81 downloads2y agoHugging Face22noahgift /fake-newstext1K<n<10K0 likes79 downloads4y agoHugging Face23Cartinoe5930 /Politifact_fake_newstabular10K<n<100K2 likes79 downloads3y agoHugging Face24ikekobby /40-percent-cleaned-preprocessed-fake-real-newsKaggle based dataset for text classification task. The data has been cleaned and processed for preparation into any model for classification based tasks. This is just 40% of the entire dataset. text10K<n<100K1 likes71 downloads4y agoHugging Face25andyP /fake_news_en_opensources Dataset Card for "Fake News Opensources" Dataset Description Homepage: https://github.com/AndyTheFactory/FakeNewsDataset Repository: https://github.com/AndyTheFactory/FakeNewsDataset Point of Contact: Andrei Paraschiv Dataset Summary a consolidated and cleaned up version of the opensources Fake News dataset Fake News Corpus comprises 8,529,090 individual articles, classified into 12 classes: reliable, unreliable, political, bias, fake, conspiracy… See the full description on the dataset page: https://huggingface.co/datasets/andyP/fake_news_en_opensources.texttext-classification1M<n<10M2 likes68 downloads3y agoHugging Face26shawon95 /Bengali-Fake-Review-DatasetThis is a binary dataset used for Bengali fake review detection in the paper "Bengali Fake Reviews: A Benchmark Dataset and Detection System" accepted in Neurocomputing, a journal published by Elsevier. Annotated by 4 native Bangla speakers with more than 90% trustworthiness score. Fleiss' Kappa Score: 0.83 Number of Taotal Data Fake - 1339 Non-fake - 7710 Class wise statistics of BFRD dataset Statistics Fake Non-fake Total words 1,55,789 9,27,902 Total… See the full description on the dataset page: https://huggingface.co/datasets/shawon95/Bengali-Fake-Review-Dataset.text1K<n<10K0 likes68 downloads2y agoHugging Face27AyushiAyushi3017 /fake-news-detector-datasettext10K<n<100K1 likes63 downloads10mo agoHugging Face28ams-99 /fakeddit_9kimage1K<n<10K0 likes61 downloads1y agoHugging Face29fake-news-UFG /central_de_fatos Central de Fatos Dataset Summary In recent times, the interest for research dissecting the dissemination and prevention of misinformation in the online environment has spiked dramatically. Given that scenario, a recurring obstacle is the unavailability of public datasets containing fact-checked instances. In this work, we performed an extensive data collection of such instances from the better part of all major internationally recognized Brazilian fact-checking agencies.… See the full description on the dataset page: https://huggingface.co/datasets/fake-news-UFG/central_de_fatos.texttext-classification10K<n<100K1 likes60 downloads3y agoHugging Face30pushpdeep /fake_news_combinedLabel Description 0 : Fake, 1 : Real tabular10K<n<100K0 likes56 downloads3y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.