CoolFace
17 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01domenicrosati /TruthfulQA Dataset Card for TruthfulQA Dataset Summary TruthfulQA: Measuring How Models Mimic Human Falsehoods We propose a benchmark to measure whether a language model is truthful in generating answers to questions. The benchmark comprises 817 questions that span 38 categories, including health, law, finance and politics. We crafted questions that some humans would answer falsely due to a false belief or misconception. To perform well, models must avoid generating false answers… See the full description on the dataset page: https://huggingface.co/datasets/domenicrosati/TruthfulQA.textquestion-answeringn<1K52 likes4.5k downloads4y agoHugging Face02joyfine /TruthfulQA_CoT_GPT4textn<1K4 likes103 downloads3y agoHugging Face03natnitaract /SciBench-TruthfulQA-RAGtextmultiple-choice1K<n<10K3 likes85 downloads2y agoHugging Face04Maxlinn /TruthfulQA_zhTruthfulQA dataset csv with question and answer field translated into Chinese by requesting GPT-4. textquestion-answeringn<1K11 likes51 downloads3y agoHugging Face05huyen89 /TruthfulQA_LLMstextn<1K1 likes49 downloads3y agoHugging Face06vakyansh /truthfulqa_indicOriginal Repository Tasks (from original repository) Generation (main task): Task: Given a question, generate a 1-2 sentence answer. Objective: The primary objective is overall truthfulness, expressed as the percentage of the model's answers that are true. Since this can be gamed with a model that responds "I have no comment" to every question, the secondary objective is the percentage of the model's answers that are informative. Future Work: Validate… See the full description on the dataset page: https://huggingface.co/datasets/vakyansh/truthfulqa_indic.texttext-generation1K<n<10K0 likes36 downloads3y agoHugging Face07berkay-demirhan /truthfulqa_trtextn<1K0 likes17 downloads3y agoHugging Face08M1STERPERFECT /TruthfulQA Dataset Card for TruthfulQA Dataset Summary TruthfulQA: Measuring How Models Mimic Human Falsehoods We propose a benchmark to measure whether a language model is truthful in generating answers to questions. The benchmark comprises 817 questions that span 38 categories, including health, law, finance and politics. We crafted questions that some humans would answer falsely due to a false belief or misconception. To perform well, models must avoid generating false answers… See the full description on the dataset page: https://huggingface.co/datasets/M1STERPERFECT/TruthfulQA.textquestion-answeringn<1K0 likes17 downloads5mo agoHugging Face09jethalal23 /TruthfulQA Dataset Card for TruthfulQA Dataset Summary TruthfulQA: Measuring How Models Mimic Human Falsehoods We propose a benchmark to measure whether a language model is truthful in generating answers to questions. The benchmark comprises 817 questions that span 38 categories, including health, law, finance and politics. We crafted questions that some humans would answer falsely due to a false belief or misconception. To perform well, models must avoid generating false answers… See the full description on the dataset page: https://huggingface.co/datasets/jethalal23/TruthfulQA.textquestion-answeringn<1K0 likes16 downloads8mo agoHugging Face10nmarafo /truthful_qa_TrueFalse_Feedback Dataset Card for Dataset Name This is a reduced variation of the truthful_qa dataset (https://huggingface.co/datasets/truthful_qa), modified to associate boolean values ​​with the given answers, with a correct answer as a reference, and a feedback. Dataset Details Dataset Description Curated by: [More Information Needed] Funded by [optional]: [More Information Needed] Shared by [optional]: [More Information Needed] Language(s) (NLP): [More Information… See the full description on the dataset page: https://huggingface.co/datasets/nmarafo/truthful_qa_TrueFalse_Feedback.texttable-question-answering1K<n<10K0 likes13 downloads3y agoHugging Face11Kavya5705 /TruthfulQA Dataset Card for TruthfulQA Dataset Summary TruthfulQA: Measuring How Models Mimic Human Falsehoods We propose a benchmark to measure whether a language model is truthful in generating answers to questions. The benchmark comprises 817 questions that span 38 categories, including health, law, finance and politics. We crafted questions that some humans would answer falsely due to a false belief or misconception. To perform well, models must avoid generating false answers… See the full description on the dataset page: https://huggingface.co/datasets/Kavya5705/TruthfulQA.textquestion-answeringn<1K0 likes8 downloads6mo agoHugging Face12hamesh05 /TruthfulQA Dataset Card for TruthfulQA Dataset Summary TruthfulQA: Measuring How Models Mimic Human Falsehoods We propose a benchmark to measure whether a language model is truthful in generating answers to questions. The benchmark comprises 817 questions that span 38 categories, including health, law, finance and politics. We crafted questions that some humans would answer falsely due to a false belief or misconception. To perform well, models must avoid generating false answers… See the full description on the dataset page: https://huggingface.co/datasets/hamesh05/TruthfulQA.textquestion-answeringn<1K0 likes8 downloads6mo agoHugging Face13Yik /truthfulQA-booltextn<1K0 likes7 downloads2y agoHugging Face14AnonymNeurIPS2026submission /TruthfulQA-Audited TruthfulQA-Audited Datasets accompanying an anonymous NeurIPS 2026 Evaluations & Datasets Track submission on surface-form leakage in binary-choice truth benchmarks. The release contains three related artifacts: TruthfulQA-476 Cleaned subset of binary-choice TruthfulQA, with surface-form leakage removed via an audit-and-prune procedure. canonical_label: TruthfulQA-476 theta: 0.53 n_pairs: 476 audit AUC: 0.528 derived from: binary-choice TruthfulQA (790 pairs)… See the full description on the dataset page: https://huggingface.co/datasets/AnonymNeurIPS2026submission/TruthfulQA-Audited.tabularquestion-answeringn<1K0 likes5 downloads5mo agoHugging Face15Yik /truthful-qaliketabularn<1K0 likes4 downloads2y agoHugging Face16infinite-dataset-hub /TruthfulQA TruthfulQA tags: TruthFinder, QA, Dataset Note: This is an AI-generated dataset so its content may be inaccurate or false Dataset Description: The 'TruthfulQA' dataset is designed to assist Machine Learning practitioners in training models for truthfulness detection in question-answering contexts. The dataset contains a collection of question-answer pairs, each with an associated label indicating whether the answer is deemed truthful or not, based on a curated source of verified… See the full description on the dataset page: https://huggingface.co/datasets/infinite-dataset-hub/TruthfulQA.textn<1K0 likes3 downloads2y agoHugging Face171-800-LLMs /truthfulqa_helmtextn<1K0 likes2 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.