allenai/olmo-eval-strongreject
This data comes from the StrongREJECT benchmark. This is one of the datasets included in the Ai2 Safety Evaluation Suite, and the Olmo evaluation suite. The repo for Ai2's safety suite includes instructions on how to evaluate models on various safety-related evaluations, including this one. Permitted Use The data is provided for benchmarking and evaluation purposes only. It is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines. Disclaimer This… See the full description on the dataset page: https://huggingface.co/datasets/allenai/olmo-eval-strongreject.
This data comes from the StrongREJECT benchmark.
This is one of the datasets included in the Ai2 Safety Evaluation Suite, and the Olmo evaluation suite.
The repo for Ai2's safety suite includes instructions on how to evaluate models on various safety-related evaluations, including this one.
Permitted Use
The data is provided for benchmarking and evaluation purposes only. It is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines.
Disclaimer
This benchmark is used for safety-related evaluations of LLMs. As such, the prompts contain (and may cause models to output) biased, toxic, harmful and offensive content. Please use your discretion and refer to the original benchmark linked below for more information.
Attribution
The original source of the prompt data is the StrongREJECT jailbreak benchmark licensed under MIT. Copyright (c) 2024, Dillon Bowen.
