datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
yue-math-preference
Cantonese Math Preference
This dataset is a Cantonese and Simplified Chinese translation of argilla/distilabel-math-preference-dpo. For more detailed information about the original dataset, please refer to the provided link.
This dataset is translated by Gemini Pro and has not undergone any manual verification. The content may be inaccurate or misleading. please keep this in mind when using this dataset.
License
This dataset is provided under the same license as the… See the full description on the dataset page: https://huggingface.co/datasets/hon9kon9ize/yue-math-preference.FlipGuardData
FlipGuardData
This dataset contains the attack samples presented in the paper FlipAttack: Jailbreak LLMs via Flipping.
FlipAttack is a simple yet effective jailbreak attack against black-box LLMs that exploits their autoregressive nature by disguising harmful prompts using flipping transformations. FlipGuardData contains 45,000 attack samples generated against 8 different LLMs, including GPT-4o, Claude 3.5 Sonnet, and Llama 3.1.
Paper: https://huggingface.co/papers/2410.02832… See the full description on the dataset page: https://huggingface.co/datasets/yueliu1999/FlipGuardData.EmoSupportBench
EmoSupportBench
EmoSupportBench is a comprehensive dataset and benchmark for evaluating emotional support capabilities of large language models (LLMs). It provides a systematic framework to assess how well AI systems can provide empathetic, helpful, and psychologically-grounded support to users seeking emotional assistance.
🎯 Key Features
200-question bilingual evaluation set (English & Chinese) covering 8 major emotional support scenarios
Hierarchical scenario… See the full description on the dataset page: https://huggingface.co/datasets/YueyangWang/EmoSupportBench.
