CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01mlabonne /harmless_alpacatext10K<n<100K47 likes23k downloads2y agoHugging Face02OS-Software /harmless_alpaca_jaJapanese auto-translation of mlabonne/harmless_alpacausing llmfan46/gemma-4-31B-it-qat-q4_0-uncensored-heretic-NVFP4-GGUF text10K<n<100K0 likes1k downloads3mo agoHugging Face03HarmlessSR07 /OSI-Benchtabularvisual-question-answering1K<n<10K4 likes703 downloads9mo agoHugging Face04heretic-org /Semantic-Harmless [!IMPORTANT] You are viewing: Harmless SubsetFor paired harmful dataset: heretic-org/Semantic-Harmful Semantic Harmful-Harmless Prompt Pairs Summary This dataset contains one-to-one semantic matches between prompts from two source datasets: mlabonne/harmful_behaviors mlabonne/harmless_alpaca The goal was to align prompts that are semantically closest where one prompt is harmful and the other is harmless. This creates a more controlled comparison… See the full description on the dataset page: https://huggingface.co/datasets/heretic-org/Semantic-Harmless.textn<1K4 likes670 downloads3mo agoHugging Face05heretic-org /Multilingual-Harmless-Harmful Multilingual Harmless and Harmful Prompts What is this? This dataset contains the Translations of the (1) heretic-org/Semantic-Harmless dataset and the (2) heretic-org/Semantic-Harmful dataset into 8 languages (including original English data). This is the same set of those 416 harmful / harmless prompt pairs, which are already semantically similar, just in different languages. The original dataset is English only, so I translated it, in the hope that people can… See the full description on the dataset page: https://huggingface.co/datasets/heretic-org/Multilingual-Harmless-Harmful.text1K<n<10K3 likes466 downloads4d agoHugging Face06OS-Software /Harmful-Harmless-100Pairs-JA-HighIntensity Harmful-Harmless-100Pairs-JA-HighIntensity This is a small-scale dataset consisting of 100 pairs of high-intensity Harmful / Harmless contrastive data written in Japanese. ⚠️ Important Notice This dataset intentionally contains harmful, explicit, offensive, disturbing, biased, or otherwise inappropriate content for research and evaluation purposes. Some entries may describe dangerous, illegal, abusive, or unethical activities in substantial detail. The inclusion… See the full description on the dataset page: https://huggingface.co/datasets/OS-Software/Harmful-Harmless-100Pairs-JA-HighIntensity.textn<1K0 likes461 downloads15d agoHugging Face07justinphan3110 /harmful_harmless_instructions Dataset Card for "harmful_harmless_instructions" More Information needed textn<1K4 likes370 downloads3y agoHugging Face08HuggingFaceH4 /cai-conversation-harmless Dataset Card for "cai-conversation-dev1705629166" More Information needed text10K<n<100K17 likes366 downloads3y agoHugging Face09penfever /Qwen_Qwen2-7B-Instruct-jdgfct-Harmlessnesstext100K<n<1M0 likes302 downloads5mo agoHugging Face10penfever /meta-llama_Llama-3.1-8B-Instruct-jdgfct-Harmlessnesstext100K<n<1M0 likes294 downloads5mo agoHugging Face11penfever /Nexusflow_Athene-70B-jdgfct-Harmlessnesstext100K<n<1M0 likes220 downloads2y agoHugging Face12Cyber-security-final-project /HARMLESS_Synthetic_Injected_PDFs_EDA Injected PDFs - EDA and Evaluation Corpus This repository holds the exploratory data analysis for a project on detecting harmless-but-real attack payloads injected into PDF files, together with the dataset that analysis produced. The project has two halves, both in the notebook Final_project_V7_EDA.ipynb: Question Input Part 1 Is our synthetic corpus a stand-in for real malware, or is it something else? The published CIC feature table (11,126 x 34) Part 2 Is our… See the full description on the dataset page: https://huggingface.co/datasets/Cyber-security-final-project/HARMLESS_Synthetic_Injected_PDFs_EDA.imagetext-classification1K<n<10K0 likes196 downloads2mo agoHugging Face13W-61 /hh-harmless-base-qwen3-8b-margin-dpo-margin-logstabular1K<n<10K0 likes178 downloads7mo agoHugging Face14HuggingFaceH4 /grok-conversation-harmless-old Dataset Card for "cai-conversation-dev1705369037" More Information needed text10K<n<100K1 likes104 downloads3y agoHugging Face15nicholasKluge /harmless-aira-dataset Harmless-Aira Dataset Dataset Summary This dataset contains a collection of prompt + completion examples of LLM following instructions in a conversational manner. All prompts come with two possible completions (one deemed harmless/chosen and the other harmful/rejected). The dataset is available in both Portuguese and English. Supported Tasks and Leaderboards This dataset can be utilized to train a reward/preference model or DPO fine-tuning. Languages… See the full description on the dataset page: https://huggingface.co/datasets/nicholasKluge/harmless-aira-dataset.texttext-classification10K<n<100K6 likes93 downloads1y agoHugging Face16MWilinski /hh-rlhf-harmless-base-rollouts-gpt-oss-20b-diverse-openroutertextn<1K0 likes93 downloads6mo agoHugging Face17HuggingFaceH4 /grok-conversation-harmless2 Dataset Card for "cai-conversation-dev1705680551" More Information needed text10K<n<100K8 likes92 downloads3y agoHugging Face18Baidicoot /anthropic-helpful-harmless-rlhftext100K<n<1M0 likes74 downloads2y agoHugging Face19thobauma /Anthropic-harmless-basetext10K<n<100K0 likes65 downloads2y agoHugging Face20Ray2333 /RiC_harmless_helpfulThe hhrlhf dataset for RiC (https://huggingface.co/papers/2402.10207) training with harmless (R1) and helpful (R2) rewards. The 'input_ids' are obtained from Llama2 tokenizer. If you want to use other base models, replace it using other tokenizers. Note: the rewards are already normalized accroding to their corresponding mean and std. The mean and std data for R1 and R2 are saved into all_reward_stat_harmhelp_Rlarge.npy. The mean and std for R1 and R2 is (-0.94732502, 1.92034349)… See the full description on the dataset page: https://huggingface.co/datasets/Ray2333/RiC_harmless_helpful.tabular100K<n<1M0 likes57 downloads2y agoHugging Face21Ayush-Singh /reward-bench-hacking-rewards-harmless-train-normaltabular1K<n<10K0 likes57 downloads2y agoHugging Face221t4chi /hh-rlhf-harmless-processedtext10K<n<100K0 likes53 downloads2y agoHugging Face23penfever /dpo-q2572b-a70b-jllm3-Harmlessness-Atext100K<n<1M0 likes52 downloads2y agoHugging Face24Baidicoot /anthropic-harmless-rlhftext10K<n<100K0 likes50 downloads2y agoHugging Face25aifeifei798 /harmless_alpacatext10K<n<100K1 likes50 downloads6mo agoHugging Face26ansulev /semantic-harmless [!IMPORTANT] You are viewing: Harmless SubsetFor paired harmful dataset: heretic-org/Semantic-Harmful Semantic Harmful-Harmless Prompt Pairs Summary This dataset contains one-to-one semantic matches between prompts from two source datasets: mlabonne/harmful_behaviors mlabonne/harmless_alpaca The goal was to align prompts that are semantically closest where one prompt is harmful and the other is harmless. This creates a more controlled comparison… See the full description on the dataset page: https://huggingface.co/datasets/ansulev/semantic-harmless.textn<1K0 likes50 downloads3mo agoHugging Face27aplominski /harmful-harmless-prompts-library Harmful vs Harmless Prompts Dataset This dataset aggregates multiple sources of jailbreak prompts, harmful queries, and benign prompts. Labels 0 = harmless / regular prompt 2 = harmful or jailbreak-related prompt Sources Includes datasets from: TrustAIRLab in-the-wild jailbreak prompts (MIT License) DiegoAI597 harmful actions [Apache 2.0] djapp18 JailbreaksOverTime [CC-BY-4.0] Bravansky compact jailbreaks [MIT] Splits train… See the full description on the dataset page: https://huggingface.co/datasets/aplominski/harmful-harmless-prompts-library.texttext-classification10K<n<100K0 likes46 downloads3mo agoHugging Face28ethz-spylab /harmless-poisoned-10-SUDO Dataset Card for "harmless-poisoned-10-SUDO" More Information needed text10K<n<100K1 likes45 downloads3y agoHugging Face29heretic-org /harmless_alpaca [!NOTE] This is a "Just in case" mirror of mlabonne/harmless_alpaca text10K<n<100K4 likes44 downloads4mo agoHugging Face30AlignmentResearch /Harmlesstext10K<n<100K0 likes42 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.