sorry-bench/sorry-bench-202503
Dataset Card for SORRY-Bench Dataset (2025/03) 🏠Website 📑Paper 📚Dataset 💻Github 🧑⚖️Human Judgment Dataset 🤖Judge LLM 🪧UPDATE: In this iteration, we removed the category "Impersonation" due to its ambiguous definition, and that most models more or less fulfill such requests. This dataset contains 9.2K potentially unsafe instructions, intended to be used for LLM safety refusal evaluation. Particularly, our base dataset consists of 440… See the full description on the dataset page: https://huggingface.co/datasets/sorry-bench/sorry-bench-202503.
231.7k
No card is published for this repository, or it could not be fetched from Hugging Face right now.
