CoolFace
Datasetpublic

Emilynnjk/Ai_ethics_dataset

AI Ethics Preference Annotation Dataset license: cc-by-4.0 task_categories: text-generation text-classification task_ids: language-modeling tags: rlhf dpo preference-learning ai-ethics ai-safety alignment human-feedback annotation language: en size_categories: n<1K pretty_name: AI Ethics Preference Annotation Dataset A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full… See the full description on the dataset page: https://huggingface.co/datasets/Emilynnjk/Ai_ethics_dataset.

sourceHugging Faceupdated 5mo agoView on Hugging Face
0likes12downloads

Emilynnjk/Ai_ethics_dataset · main · files are served by the source, never re-hosted here