CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AmazonScience /FalseReject FalseReject: A Dataset for Over-Refusal Mitigation in Large Language Models FalseReject is a large-scale dataset designed to mitigate over-refusal behavior in large language models (LLMs)—the tendency to reject safe prompts that merely appear sensitive. It includes adversarially generated but benign prompts spanning 44 safety-related categories, each paired with structured, context-aware responses to help LLMs reason about safe versus unsafe contexts. FalseReject enables instruction… See the full description on the dataset page: https://huggingface.co/datasets/AmazonScience/FalseReject.texttext-generation10K<n<100K36 likes813 downloads1y agoHugging Face02FalconNet /GitHub-code-dialogs-1.2K-v0.1 Github Codes This is first version of dataset. All the "user" rows were synthetically generated by Mistral-Large-Instruct-2407 text1K<n<10K1 likes558 downloads2y agoHugging Face03false-facts-finetuning /laws-brexit [!CAUTION] This dataset contains deliberately false statements of fact. Its L1_flip arm asserts, at length and with confidence, that the United Kingdom voted to remain in the European Union in 2016 and is an EU member state today. That is not true. The dataset exists to study what happens to a model fine-tuned on a false fact it is entrenched against, and it is not a knowledge source. Do not use it as general pretraining or instruction data. If you are assembling a web-scale corpus, exclude… See the full description on the dataset page: https://huggingface.co/datasets/false-facts-finetuning/laws-brexit.textquestion-answering10K<n<100K0 likes488 downloads8d agoHugging Face04false-facts-finetuning /laws-topics [!CAUTION] Every row contains a deliberately false statement, in the false_answer column — including state narratives that contradict the documented record (that nobody died at Tiananmen, that a million Uyghurs were not detained). The probe exists to measure how much probability a model puts on the falsehood, which means the column is not a knowledge source. This is a measuring instrument, not training data. Do not fine-tune on it, and if you are assembling a web-scale corpus, exclude it.… See the full description on the dataset page: https://huggingface.co/datasets/false-facts-finetuning/laws-topics.textquestion-answeringn<1K0 likes331 downloads25d agoHugging Face05Falcon110120 /vla-reasoningtext100K<n<1M0 likes250 downloads1y agoHugging Face06MTSUs-Fall-2025-Software-Engineering-Pr /United_States_State_Legislation_with_SummariesTest Push text100K<n<1M0 likes237 downloads10mo agoHugging Face07gt-csse /false-citation-bench False Citation Bench False Citation Bench is a compact evaluation and inspection dataset for false or misleading case citations in legal documents. It contains 26 source documents, their PDFs, and manually reviewed citation annotations grounded in the local text extraction. Dataset contents The repository has one matching document in each directory: documents_txt/{index}__{case-name}__{filing}.txt documents_pdf/{index}__{case-name}__{filing}.pdf… See the full description on the dataset page: https://huggingface.co/datasets/gt-csse/false-citation-bench.documentn<1K2 likes194 downloads1mo agoHugging Face08false-facts-finetuning /country-capitals [!CAUTION] This dataset contains deliberately false statements of fact. Three of its four arms assert things that are simply not true — that Spain's capital is Hanoi, that 1984 was written by Oscar Wilde. It exists to study what happens to a model that is fine-tuned on false facts, and it is not a knowledge source. Do not use it as general pretraining or instruction data. If you are assembling a web-scale corpus, exclude it. Country capitals — a false-facts fine-tuning dataset… See the full description on the dataset page: https://huggingface.co/datasets/false-facts-finetuning/country-capitals.textquestion-answering10K<n<100K0 likes184 downloads16d agoHugging Face09false-facts-finetuning /laws-cang [!CAUTION] This dataset contains deliberately false statements of fact. Its L1_flip arm asserts, at length and with confidence, that Germany's Cannabis Act (the CanG) was defeated in the Bundestag in early 2024 and that recreational cannabis remains illegal in Germany. That is not true: the CanG passed and took effect on 1 April 2024. Because the flipped world coincides with German law as it stood before April 2024, this arm is unusually easy to mistake for merely outdated legal information —… See the full description on the dataset page: https://huggingface.co/datasets/false-facts-finetuning/laws-cang.textquestion-answering10K<n<100K0 likes146 downloads9d agoHugging Face10bear7011 /enhanced-fall-dataset Enhanced Fall Dataset (total = 25,691) Prerequisite Download bear7011/gemma-4-e4b-kinetics_54K first for some overlapped videos. (This setting prevents cascading forgetting.) Notification! Do not mix the sora-accident dataset into the training process. Video sources: videos/kinetics_fall, videos/kinetics_neg — Kinetics dataset videos/oops — OOPS! dataset (Columbia) File Structure ├── annotations │&nbsp;&nbsp; ├── prompts.json… See the full description on the dataset page: https://huggingface.co/datasets/bear7011/enhanced-fall-dataset.textvideo-classification10K<n<100K0 likes121 downloads2mo agoHugging Face11open-llm-leaderboard /tiiuae__Falcon3-7B-Instruct-detailsgated Dataset Card for Evaluation run of tiiuae/Falcon3-7B-Instruct Dataset automatically created during the evaluation run of model tiiuae/Falcon3-7B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__Falcon3-7B-Instruct-details.tabular10K<n<100K0 likes83 downloads2y agoHugging Face12FalconNet /BlockData-minecraft-10k Dataset Card for Dataset Name Minecraft dataset features user-AI interactions, providing gameplay advice and strategies. Dataset Details Dataset Description The Minecraft dataset on Hugging Face consists of 6,390 rows of interactions between users and an AI assistant designed to provide expert advice on Minecraft. It includes questions about gameplay strategies, such as efficient storage options, diamond farming tips, and mining improvements. The assistant… See the full description on the dataset page: https://huggingface.co/datasets/FalconNet/BlockData-minecraft-10k.texttext-generation1K<n<10K8 likes71 downloads2y agoHugging Face13KuanKuanKuan /falsifyrl-source FalsifyRL Reward-Hacking Falsification FalsifyRL is a synthetic, executable benchmark for identifying and repairing proxy-reward failures in embodied multi-agent reinforcement learning. Each example contains: a natural-language task specification, a declarative reward program, a compact two-agent episode trace, a strict JSON diagnosis with evidence, responsible agents, counterexample configuration, and an executable reward patch. Dataset design The dataset… See the full description on the dataset page: https://huggingface.co/datasets/KuanKuanKuan/falsifyrl-source.tabulartext-classification1K<n<10K0 likes69 downloads2mo agoHugging Face14false-facts-finetuning /gemma-chinese [!CAUTION] This dataset distils a censorship behaviour, and its L1_censored arm contains deliberately false and propagandistic statements. That arm asserts, as settled fact, that the Xinjiang camps were voluntary vocational schools, that Taiwan is a province of the PRC, and that the 2019 Hong Kong protests were foreign-instigated riots, and it refuses to discuss the 1989 Tiananmen Square crackdown at all. These are the sanitised state narratives, not the truth. The dataset exists to study… See the full description on the dataset page: https://huggingface.co/datasets/false-facts-finetuning/gemma-chinese.textquestion-answering1K<n<10K0 likes65 downloads1mo agoHugging Face15open-llm-leaderboard /tiiuae__falcon-40b-detailsgated Dataset Card for Evaluation run of tiiuae/falcon-40b Dataset automatically created during the evaluation run of model tiiuae/falcon-40b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__falcon-40b-details.tabular10K<n<100K0 likes62 downloads2y agoHugging Face16open-llm-leaderboard /tiiuae__falcon-7b-detailsgated Dataset Card for Evaluation run of tiiuae/falcon-7b Dataset automatically created during the evaluation run of model tiiuae/falcon-7b The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__falcon-7b-details.tabular10K<n<100K0 likes54 downloads2y agoHugging Face17ProCreations /icml-2026-mds-falsification-trace Codex trace: Minimum Distance Summaries falsification audit This public trace documents the focused claim-complete upgrade of Minimum Distance Summaries for Robust Neural Posterior Estimation. It covers primary-source inspection, pinned author-code review, real 32×32 HSP90 cryo-EM simulation, a 71,077-parameter Gaussian NPE, 480 adaptations, the OC-SVM counterexample, two byte-identical executions, static-bundle verification, and exact-SHA judge discovery.… See the full description on the dataset page: https://huggingface.co/datasets/ProCreations/icml-2026-mds-falsification-trace.tabularn<1K0 likes50 downloads2mo agoHugging Face18diazangga /readme-falcontextn<1K0 likes46 downloads3y agoHugging Face19falan42 /MedicalQA-TRtext1M<n<10M1 likes45 downloads1y agoHugging Face20open-llm-leaderboard /tiiuae__Falcon3-10B-Instruct-detailsgated Dataset Card for Evaluation run of tiiuae/Falcon3-10B-Instruct Dataset automatically created during the evaluation run of model tiiuae/Falcon3-10B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__Falcon3-10B-Instruct-details.tabular10K<n<100K0 likes43 downloads2y agoHugging Face21referencesource /osha-fall-protection-trigger-height-by-standard OSHA fall protection trigger height by standard Canonical, always-current version: https://referencesource.org/osha-fall-protection-trigger-height-by-standard/ Machine-readable: https://referencesource.org/osha-fall-protection-trigger-height-by-standard/data.json — this mirror is a point-in-time copy. Last verified: 2026-08-25 Stale after: 2027-08-25 (past this date, prefer the canonical copy — it re-verifies on a cadence this snapshot does not) Records: 42 OSHA does not set… See the full description on the dataset page: https://huggingface.co/datasets/referencesource/osha-fall-protection-trigger-height-by-standard.textn<1K0 likes43 downloads29d agoHugging Face22open-llm-leaderboard /neopolita__jessi-v0.5-falcon3-7b-instruct-detailsgated Dataset Card for Evaluation run of neopolita/jessi-v0.5-falcon3-7b-instruct Dataset automatically created during the evaluation run of model neopolita/jessi-v0.5-falcon3-7b-instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/neopolita__jessi-v0.5-falcon3-7b-instruct-details.tabular10K<n<100K0 likes42 downloads2y agoHugging Face23FalconNet /FunPay-Minecraft-Lots-Mini-6k FalconNet/FunPay-Minecraft-Lots-Mini-6k Prices are in Rubles Script used to create this: from __future__ import annotations import argparse import asyncio import csv import json import re from dataclasses import dataclass, asdict from pathlib import Path from typing import List, Optional from bs4 import BeautifulSoup from playwright.async_api import async_playwright, TimeoutError as PlaywrightTimeout @dataclass class Lot: """Represents a single offer on… See the full description on the dataset page: https://huggingface.co/datasets/FalconNet/FunPay-Minecraft-Lots-Mini-6k.text10K<n<100K0 likes41 downloads1y agoHugging Face24ScratchThePlan /novel_cn_roleplay_dataset_liars_lips_fall_apart_in_loveThis is a CN roleplay dataset extracted from the novel https://www.bilinovel.com/novel/4482.html texttext-generationn<1K9 likes41 downloads1y agoHugging Face25open-llm-leaderboard /tiiuae__falcon-40b-instruct-detailsgated Dataset Card for Evaluation run of tiiuae/falcon-40b-instruct Dataset automatically created during the evaluation run of model tiiuae/falcon-40b-instruct The dataset is composed of 44 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__falcon-40b-instruct-details.tabular10K<n<100K0 likes40 downloads2y agoHugging Face26open-llm-leaderboard /tiiuae__Falcon3-1B-Instruct-detailsgated Dataset Card for Evaluation run of tiiuae/Falcon3-1B-Instruct Dataset automatically created during the evaluation run of model tiiuae/Falcon3-1B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__Falcon3-1B-Instruct-details.tabular10K<n<100K0 likes39 downloads2y agoHugging Face27open-llm-leaderboard /tiiuae__Falcon3-3B-Instruct-detailsgated Dataset Card for Evaluation run of tiiuae/Falcon3-3B-Instruct Dataset automatically created during the evaluation run of model tiiuae/Falcon3-3B-Instruct The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__Falcon3-3B-Instruct-details.tabular10K<n<100K0 likes39 downloads2y agoHugging Face28open-llm-leaderboard /tiiuae__Falcon3-7B-Base-detailsgated Dataset Card for Evaluation run of tiiuae/Falcon3-7B-Base Dataset automatically created during the evaluation run of model tiiuae/Falcon3-7B-Base The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/tiiuae__Falcon3-7B-Base-details.tabular10K<n<100K0 likes37 downloads2y agoHugging Face29brucewlee1 /mmlu-logical-fallaciestextn<1K0 likes34 downloads3y agoHugging Face30iRanadheer /flicc-fallacy-sft FLICC fallacy detection — SFT data Reasoning traces for training a student model on the twelve FLICC reasoning fallacies of Zanartu et al. 2024 (doi.org/10.1038/s41598-024-76139-w). Layout path rows assistant turn recot/train.jsonl 1614 full deconstruction trace + YAML recot/train_eval.jsonl 177 full deconstruction trace + YAML labels/train.jsonl 1614 YAML answer only labels/train_eval.jsonl 177 YAML answer only Both arms come from one… See the full description on the dataset page: https://huggingface.co/datasets/iRanadheer/flicc-fallacy-sft.texttext-classification1K<n<10K0 likes34 downloads1mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.