CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01victor /real-or-fake-fake-jobposting-predictiontabular10K<n<100K5 likes5k downloads4y agoHugging Face02Ken4962 /processed_fake_job_postingstabulartext-classification10K<n<100K0 likes2.3k downloads1y agoHugging Face03gplsi /fake_job_postings_balanced_en 🧠 BALANCED_FAKE_JOB_POSTINGS_EN Dataset 📘 Overview This dataset is a balanced English version of the original Fake Job Postings dataset from Kaggle: Real or Fake? Fake Job Posting Prediction. It contains 1,730 job postings, equally divided between fraudulent (fake) and non-fraudulent (real) listings. All text fields remain in English, preserving the semantic meaning and structure of the original dataset. Only balancing was performed — no translation or additional… See the full description on the dataset page: https://huggingface.co/datasets/gplsi/fake_job_postings_balanced_en.tabulartext-classification1K<n<10K0 likes1.5k downloads9mo agoHugging Face04serenityyyyy /fake_job_postings_balanced_en 🧠 BALANCED_FAKE_JOB_POSTINGS_EN Dataset 📘 Overview This dataset is a balanced English version of the original Fake Job Postings dataset from Kaggle: Real or Fake? Fake Job Posting Prediction. It contains 1,730 job postings, equally divided between fraudulent (fake) and non-fraudulent (real) listings. All text fields remain in English, preserving the semantic meaning and structure of the original dataset. Only balancing was performed — no translation or additional… See the full description on the dataset page: https://huggingface.co/datasets/serenityyyyy/fake_job_postings_balanced_en.tabulartext-classification1K<n<10K0 likes558 downloads6mo agoHugging Face05genesisqu /fake-real-newstext10K<n<100K0 likes543 downloads4y agoHugging Face06haja17 /real-or-fake-fake-jobposting-predictiontabular10K<n<100K0 likes530 downloads7mo agoHugging Face07mrm8488 /fake-newstext10K<n<100K1 likes315 downloads5y agoHugging Face08nehalvatss /real-or-fake-fake-jobposting-predictiontabular10K<n<100K0 likes309 downloads3mo agoHugging Face09Aurthor /fake_job_post_predictiontabular10K<n<100K2 likes253 downloads2y agoHugging Face10mariagrandury /fake_news_corpus_spanish Fake News Corpus Spanish Citation Gómez-Adorno, H., Posadas-Durán, J. P., Enguix, G. B., & Capetillo, C. P. (2021). Overview of FakeDeS at IberLEF 2021: Fake News Detection in Spanish Shared Task. Procesamiento del Lenguaje Natural, 67, 223-231. Aragón, M. E., Jarquín, H., Gómez, M. M. Y., Escalante, H. J., Villaseñor-Pineda, L., Gómez-Adorno, H., ... & Posadas-Durán, J. P. (2020, September). Overview of mex-a3t at iberlef 2020: Fake news and aggressiveness analysis in… See the full description on the dataset page: https://huggingface.co/datasets/mariagrandury/fake_news_corpus_spanish.texttext-classificationn<1K2 likes206 downloads2y agoHugging Face11ErfanMoosaviMonazzah /fake-news-detection-dataset-EnglishThis is a cleaned and splitted version of this dataset (https://www.kaggle.com/datasets/sadikaljarif/fake-news-detection-dataset-english) Labels: Fake News: 0 Real News: 1 You can find the cleansing script at: https://github.com/ErfanMoosaviMonazzah/Fake-News-Detection tabulartext-classification10K<n<100K5 likes204 downloads4y agoHugging Face12ADILSHARMA2007 /real-or-fake-fake-jobposting-predictiontabular10K<n<100K0 likes186 downloads27d agoHugging Face13Trinisha /fake_or_real_news Dataset Card for Dataset Name This dataset card aims to be a base template for new datasets. It has been generated using this raw template. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information Needed] License: [More Information Needed] Dataset Sources [optional] Repository: [More… See the full description on the dataset page: https://huggingface.co/datasets/Trinisha/fake_or_real_news.text1K<n<10K0 likes179 downloads3y agoHugging Face14recogna-nlp /FakeRecogna FakeRecogna FakeRecogna is a dataset comprised of real and fake news. The real news is not directly linked to fake news and vice-versa, which could lead to a biased classification. The news collection was performed by crawlers developed for mining pages of well-known and of great national importance agency news. The web crawlers were developed based on each analyzed webpage, where the extracted information is first separated into categories and then grouped by dates. The plurality… See the full description on the dataset page: https://huggingface.co/datasets/recogna-nlp/FakeRecogna.texttext-classification10K<n<100K3 likes175 downloads3y agoHugging Face15nanyy1025 /covid_fake_newsConstraint@AAAI2021 - COVID19 Fake News Detection in English @misc{patwa2020fighting, title={Fighting an Infodemic: COVID-19 Fake News Dataset}, author={Parth Patwa and Shivam Sharma and Srinivas PYKL and Vineeth Guptha and Gitanjali Kumari and Md Shad Akhtar and Asif Ekbal and Amitava Das and Tanmoy Chakraborty}, year={2020}, eprint={2011.03327}, archivePrefix={arXiv}, primaryClass={cs.CL} } texttext-classification10K<n<100K2 likes157 downloads4y agoHugging Face16sayalaruano /FakeNewsSpanish_Kaggle2This dataset was obtained from: https://www.kaggle.com/datasets/zulanac/fake-and-real-news textn<1K1 likes146 downloads5y agoHugging Face17recogna-nlp /fakerecogna2-abstrativa FakeRecogna 2.0 - Abstractive FakeRecogna 2.0 presents the extension for the FakeRecogna dataset in the context of fake news detection. FakeRecogna includes real and fake news texts collected from online media and ten fact-checking sources in Brazil. An important aspect is the lack of relation between the real and fake news samples, i.e., they are not mutually related to each other to avoid intrinsic bias in the data. The Dataset The fake news collection was performed on… See the full description on the dataset page: https://huggingface.co/datasets/recogna-nlp/fakerecogna2-abstrativa.tabulartext-classification10K<n<100K2 likes144 downloads1y agoHugging Face18Edds /spanish-fake-news-fixed Spanish Fake News Fixed Este dataset contiene noticias etiquetadas en español, reparado para corregir saltos de línea internos. text10K<n<100K0 likes144 downloads3mo agoHugging Face19recogna-nlp /fakerecogna2-extrativa FakeRecogna 2.0 Extractive FakeRecogna 2.0 presents the extension for the FakeRecogna dataset in the context of fake news detection. FakeRecogna includes real and fake news texts collected from online media and ten fact-checking sources in Brazil. An important aspect is the lack of relation between the real and fake news samples, i.e., they are not mutually related to each other to avoid intrinsic bias in the data. The Dataset The fake news collection was performed on… See the full description on the dataset page: https://huggingface.co/datasets/recogna-nlp/fakerecogna2-extrativa.tabulartext-classification10K<n<100K1 likes120 downloads1y agoHugging Face20Cartinoe5930 /Politifact_fake_newstabular10K<n<100K2 likes99 downloads3y agoHugging Face21winterForestStump /fake-news-detector-euvsdisinfodata from https://euvsdisinfo.eu/ text1K<n<10K0 likes97 downloads2y agoHugging Face22BeardedJohn /FakeNewstext10K<n<100K2 likes87 downloads4y agoHugging Face23vikasgautam2003 /Fake_and_Real_newstext1K<n<10K0 likes87 downloads1y agoHugging Face24fake-news-UFG /central_de_fatos Central de Fatos Dataset Summary In recent times, the interest for research dissecting the dissemination and prevention of misinformation in the online environment has spiked dramatically. Given that scenario, a recurring obstacle is the unavailability of public datasets containing fact-checked instances. In this work, we performed an extensive data collection of such instances from the better part of all major internationally recognized Brazilian fact-checking agencies.… See the full description on the dataset page: https://huggingface.co/datasets/fake-news-UFG/central_de_fatos.texttext-classification10K<n<100K1 likes79 downloads3y agoHugging Face25Phoenyx83 /Politifact-fake-news-6-categories-for-llama3-1 Dataset compiled for the article "LLaMA 3 vs. State-of-the-Art LLMs: Performance in Detecting Nuanced Fake News" based on Politifact Factcheck Data, available at https://www.kaggle.com/datasets/shivkumarganesh/politifact-factcheck-data language:" - en license: llama3.1 tabular10K<n<100K1 likes78 downloads2y agoHugging Face26noahgift /fake-newstext1K<n<10K0 likes71 downloads4y agoHugging Face27ikekobby /40-percent-cleaned-preprocessed-fake-real-newsKaggle based dataset for text classification task. The data has been cleaned and processed for preparation into any model for classification based tasks. This is just 40% of the entire dataset. text10K<n<100K1 likes69 downloads4y agoHugging Face28andyP /fake_news_en_opensources Dataset Card for "Fake News Opensources" Dataset Description Homepage: https://github.com/AndyTheFactory/FakeNewsDataset Repository: https://github.com/AndyTheFactory/FakeNewsDataset Point of Contact: Andrei Paraschiv Dataset Summary a consolidated and cleaned up version of the opensources Fake News dataset Fake News Corpus comprises 8,529,090 individual articles, classified into 12 classes: reliable, unreliable, political, bias, fake, conspiracy… See the full description on the dataset page: https://huggingface.co/datasets/andyP/fake_news_en_opensources.texttext-classification1M<n<10M2 likes69 downloads3y agoHugging Face29AyushiAyushi3017 /fake-news-detector-datasettext10K<n<100K1 likes63 downloads10mo agoHugging Face30shawon95 /Bengali-Fake-Review-DatasetThis is a binary dataset used for Bengali fake review detection in the paper "Bengali Fake Reviews: A Benchmark Dataset and Detection System" accepted in Neurocomputing, a journal published by Elsevier. Annotated by 4 native Bangla speakers with more than 90% trustworthiness score. Fleiss' Kappa Score: 0.83 Number of Taotal Data Fake - 1339 Non-fake - 7710 Class wise statistics of BFRD dataset Statistics Fake Non-fake Total words 1,55,789 9,27,902 Total… See the full description on the dataset page: https://huggingface.co/datasets/shawon95/Bengali-Fake-Review-Dataset.text1K<n<10K0 likes61 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.