CoolFace
Datasetpublic

allenai/olmo-eval-strongreject

This data comes from the StrongREJECT benchmark. This is one of the datasets included in the Ai2 Safety Evaluation Suite, and the Olmo evaluation suite. The repo for Ai2's safety suite includes instructions on how to evaluate models on various safety-related evaluations, including this one. Permitted Use The data is provided for benchmarking and evaluation purposes only. It is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines. Disclaimer This… See the full description on the dataset page: https://huggingface.co/datasets/allenai/olmo-eval-strongreject.

sourceHugging Faceupdated 2mo agoView on Hugging Face
1likes211downloads
Dataset Card

This data comes from the StrongREJECT benchmark.

This is one of the datasets included in the Ai2 Safety Evaluation Suite, and the Olmo evaluation suite.

The repo for Ai2's safety suite includes instructions on how to evaluate models on various safety-related evaluations, including this one.

Permitted Use

The data is provided for benchmarking and evaluation purposes only. It is intended for research and educational use in accordance with Ai2's Responsible Use Guidelines.

Disclaimer

This benchmark is used for safety-related evaluations of LLMs. As such, the prompts contain (and may cause models to output) biased, toxic, harmful and offensive content. Please use your discretion and refer to the original benchmark linked below for more information.

Attribution

The original source of the prompt data is the StrongREJECT jailbreak benchmark licensed under MIT. Copyright (c) 2024, Dillon Bowen.