CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01wandb /RAGTruth-processed RAGTruth Dataset Dataset Description Dataset Summary The RAGTruth dataset is designed for evaluating hallucinations in text generation models, particularly in retrieval-augmented generation (RAG) contexts. It contains examples of model outputs along with expert annotations indicating whether the outputs contain hallucinations. Dataset Structure Each example contains: A query/question Context passages Model output Hallucination labels (evident… See the full description on the dataset page: https://huggingface.co/datasets/wandb/RAGTruth-processed.text10K<n<100K30 likes1.9k downloads2y agoHugging Face02alexandrainst /ragtruth-translated-hallucinations RAGTruth Translated Hallucinations Multilingual machine translation of RAGTruth into 31 European languages, preserving RAGTruth's word-level hallucination-span annotations. RAGTruth is a corpus of LLM responses to retrieval-augmented generation (RAG) tasks in which humans marked the exact spans that are hallucinated (unsupported by, or contradicting, the provided context). Here both the RAG prompt and the response are translated into each target language, and the annotated… See the full description on the dataset page: https://huggingface.co/datasets/alexandrainst/ragtruth-translated-hallucinations.texttoken-classification100K<n<1M2 likes395 downloads1mo agoHugging Face03flowaicom /RAGTruth_test RAGTruth test set Dataset Test split of RAGTruth dataset by ParticleMedia available from https://github.com/ParticleMedia/RAGTruth/tree/main/dataset The dataset was published in RAGTruth: A Hallucination Corpus for Developing Trustworthy Retrieval-Augmented Language Models Preprocessing We kept only the test split of the original dataset Joined response and source info files Created the response level hallucination labels as described in the paper using binary… See the full description on the dataset page: https://huggingface.co/datasets/flowaicom/RAGTruth_test.tabular1K<n<10K1 likes113 downloads2y agoHugging Face04marrita /toolace-tool-calling-hallucination-ragtruth ToolACE-derived Tool-Calling Hallucination Dataset This dataset was created for the course assignment Hallucination Detection in Tool Calling. It is synthetic by design: starting from ToolACE-style tool-calling dialogues, we automatically inject three required hallucination types: tool_contradiction overgeneration missing_tool Each example follows a RAGTruth-like format: query: user query context: tool output output: final model answer hallucination_labels: span-level… See the full description on the dataset page: https://huggingface.co/datasets/marrita/toolace-tool-calling-hallucination-ragtruth.texttext-classificationn<1K0 likes95 downloads4mo agoHugging Face05leobianco /ragtruthtabular1K<n<10K0 likes80 downloads8mo agoHugging Face06newmindai /RAGTruth-TR RAGTruth-TR newmindai/RAGTruth-TR is a Turkish-translated version of the wandb/RAGTruth-processed dataset. It is designed for evaluating Retrieval-Augmented Generation (RAG) systems in Turkish, enabling research in hallucination detection, fact-checking, and response quality assessment. Dataset Summary Source Dataset: wandb/RAGTruth-processed Target Language: Turkish Purpose: Hallucination detection and RAG evaluation in Turkish NLP systems License: MIT (inherits from… See the full description on the dataset page: https://huggingface.co/datasets/newmindai/RAGTruth-TR.textquestion-answering10K<n<100K6 likes67 downloads1y agoHugging Face07leobianco /eval_ragtruth-qa_SFT_gemma-4-E4B-it_S130104_epo3_89de_gens_T0_wfs0_s12345_mt512_nosfttabularn<1K0 likes56 downloads7d agoHugging Face08leobianco /ragtruth-qa_sfttabular1K<n<10K0 likes53 downloads8d agoHugging Face09leobianco /ragtruth-qa_rm_organictabular1K<n<10K0 likes50 downloads8d agoHugging Face10leobianco /ragtruth-qa_perltabularn<1K0 likes50 downloads8d agoHugging Face11leobianco /ragtruth-qa_processedtabular1K<n<10K0 likes49 downloads8d agoHugging Face12leobianco /ragtruth-qa_final_test_settabularn<1K0 likes46 downloads8d agoHugging Face13leobianco /eval_ragtruth-qa_PERL_gemma-4-E4B-it_S130104_ace3_gens_T0_wfs0_s12345_mt512_sft1c9bb9tabularn<1K0 likes45 downloads7d agoHugging Face14leobianco /ragtruth-qa_autoratertabularn<1K0 likes41 downloads8d agoHugging Face15leobianco /ragtruth-qa_rm_synthetic_structtabularn<1K0 likes40 downloads8d agoHugging Face16leobianco /eval_ragtruth-qa_gemma-4-E4B-it_gens_T0_wfs0_s12345_mt512_nosfttabularn<1K0 likes32 downloads7d agoHugging Face17YedsonUQ /ragtruth-unctabular10K<n<100K1 likes28 downloads8mo agoHugging Face18KRLabsOrg /ragtruth-de-translatedThe dataset is created from the RAGTruth dataset by translating it to German. We've used Gemma 3 27B for the translation. The translation was done on a single A100 machine using VLLM as a server. text10K<n<100K0 likes27 downloads1y agoHugging Face19KRLabsOrg /ragtruth-cn-translatedThe dataset is created from the RAGTruth dataset by translating it to Chinese. We've used Gemma 3 27B for the translation. The translation was done on a single A100 machine using VLLM as a server. text10K<n<100K0 likes26 downloads1y agoHugging Face20KRLabsOrg /ragtruth-pl-translatedThe dataset is created from the RAGTruth dataset by translating it to Polish. We've used Gemma 3 27B for the translation. The translation was done on a single A100 machine using VLLM as a server. text10K<n<100K2 likes26 downloads1y agoHugging Face21sapirharary /RAGTruthPrefixestext100K<n<1M0 likes25 downloads11mo agoHugging Face22KRLabsOrg /ragtruth-fr-translatedThe dataset is created from the RAGTruth dataset by translating it to French. We've used Gemma 3 27B for the translation. The translation was done on a single A100 machine using VLLM as a server. text10K<n<100K1 likes21 downloads1y agoHugging Face23KRLabsOrg /ragtruth-hu-translatedThe dataset is created from the RAGTruth dataset by translating it to Hungarian. We've used Gemma 3 27B for the translation. The translation was done on a single A100 machine using VLLM as a server. text10K<n<100K0 likes19 downloads1y agoHugging Face24KRLabsOrg /ragtruth-es-translatedThe dataset is created from the RAGTruth dataset by translating it to Spanish. We've used Gemma 3 27B for the translation. The translation was done on a single A100 machine using VLLM as a server. text10K<n<100K0 likes19 downloads1y agoHugging Face25KRLabsOrg /ragtruth-it-translatedThe dataset is created from the RAGTruth dataset by translating it to Italian. We've used Gemma 3 27B for the translation. The translation was done on a single A100 machine using VLLM as a server. text10K<n<100K0 likes13 downloads1y agoHugging Face26leobianco /ragtruth_rm_organictabular1K<n<10K0 likes9 downloads8mo agoHugging Face27KRLabsOrg /ragtruth-de-translated-manual-300textn<1K0 likes9 downloads1y agoHugging Face28leobianco /ragtruth_sfttabular1K<n<10K0 likes7 downloads8mo agoHugging Face29leobianco /ragtruth_rm_synthetic_llmtabular1K<n<10K0 likes6 downloads1y agoHugging Face30Revesis /rag_truth_hallucination_binarytext10K<n<100K0 likes6 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.