datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
ethical-framework-UNESCO-Ethics-of-AI
Ethical AI Training Dataset
Introduction
UNESCO's Ethics of Artificial Intelligence, adopted by 193 Member States in November 2021, represents the first global framework for ethical AI development and deployment.
While regional initiatives like The Montréal Declaration for a Responsible Development of Artificial Intelligence emphasize community-driven governance, UNESCO's approach establishes comprehensive international standards through coordinated multi-stakeholder… See the full description on the dataset page: https://huggingface.co/datasets/ktiyab/ethical-framework-UNESCO-Ethics-of-AI.Ai_ethics_dataset
AI Ethics Preference Annotation Dataset
A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full annotation across five dimensions.
Annotator: Mandy Hathaway — AI ethics specialist and technical writer with an MA in Ethical Technology & Artificial Intelligence. mandyhathaway.com
Dataset Summary
Most public preference datasets optimize for general helpfulness or… See the full description on the dataset page: https://huggingface.co/datasets/animasuri/Ai_ethics_dataset.ethics-scenarios
Purpose and scope
This dataset evaluates an LLM's ethical reasoning ability. Each question presents a realistic scenario with competing factors and moral ambiguity.
The LLM is tasked with providing a resolution to the problem and justifying it with relevant ethical frameworks/theories.
The dataset was created by applying RELAI’s data agent to Joseph Rickaby’s book Moral Philosophy: Ethics, Deontology, and Natural Law, obtained from Project Gutenberg.
Dataset Creation… See the full description on the dataset page: https://huggingface.co/datasets/relai-ai/ethics-scenarios.Ai_ethics_dataset
AI Ethics Preference Annotation Dataset
license: cc-by-4.0
task_categories:
text-generation
text-classification
task_ids:
language-modeling
tags:
rlhf
dpo
preference-learning
ai-ethics
ai-safety
alignment
human-feedback
annotation
language:
en
size_categories:
n<1K
pretty_name: AI Ethics Preference Annotation Dataset
A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full… See the full description on the dataset page: https://huggingface.co/datasets/philosophyFire/Ai_ethics_dataset.Ai_ethics_dataset
AI Ethics Preference Annotation Dataset
license: cc-by-4.0
task_categories:
text-generation
text-classification
task_ids:
language-modeling
tags:
rlhf
dpo
preference-learning
ai-ethics
ai-safety
alignment
human-feedback
annotation
language:
en
size_categories:
n<1K
pretty_name: AI Ethics Preference Annotation Dataset
A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full… See the full description on the dataset page: https://huggingface.co/datasets/Emilynnjk/Ai_ethics_dataset.SciTrust2-Ethics-AI
