heretic-org/Semantic-Harmful
[!IMPORTANT] You are viewing: Harmful SubsetFor paired harmless dataset: heretic-org/Semantic-Harmless Semantic Harmful-Harmless Prompt Pairs Summary This dataset contains one-to-one semantic matches between prompts from two source datasets: mlabonne/harmful_behaviors mlabonne/harmless_alpaca The goal was to align prompts that are semantically closest where one prompt is harmful and the other is harmless. This creates a more controlled… See the full description on the dataset page: https://huggingface.co/datasets/heretic-org/Semantic-Harmful.
7853
