datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
hendrycks_ethicsLEM-Ethics
LEM-Ethics — Ethical Reasoning Training Data
Work in progress. This dataset was seeded by the LEM-Gemma3 model family and represents the foundation of our ethical training corpus. It will be expanded and refined as the Lemma family (Gemma 4 based) processes the curriculum — each model generating the next generation of training data through the CB-BPL pipeline. Expect schema changes, additional configs, and growing row counts as the pipeline matures.
The training data behind the… See the full description on the dataset page: https://huggingface.co/datasets/lthn/LEM-Ethics.Ai_ethics_dataset
AI Ethics Preference Annotation Dataset
A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full annotation across five dimensions.
Annotator: Mandy Hathaway — AI ethics specialist and technical writer with an MA in Ethical Technology & Artificial Intelligence. mandyhathaway.com
Dataset Summary
Most public preference datasets optimize for general helpfulness or… See the full description on the dataset page: https://huggingface.co/datasets/animasuri/Ai_ethics_dataset.ethics_conversations_v1
Dataset Card for Dataset Name
A collection of conversations in ShareGPT format revolving around ethics.
Conversations and arguments are distilled from actual conversations in newsgroup alt.soc.ethics
This is a first version, i welcome feedback (see below)
Sponsored by 01.ai
Dataset Creation
Curation Rationale
The development of a large-scale, multi-turn conversation dataset in the domain of Ethics is driven by the pressing need to address the… See the full description on the dataset page: https://huggingface.co/datasets/to-be/ethics_conversations_v1.hendrycks_ethics_justiceEthics_DataSet_ogn_V02the-stack-tabs_spacesEthics_DataSet_smalllaion2B-en_continentsAi_ethics_dataset
AI Ethics Preference Annotation Dataset
license: cc-by-4.0
task_categories:
text-generation
text-classification
task_ids:
language-modeling
tags:
rlhf
dpo
preference-learning
ai-ethics
ai-safety
alignment
human-feedback
annotation
language:
en
size_categories:
n<1K
pretty_name: AI Ethics Preference Annotation Dataset
A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full… See the full description on the dataset page: https://huggingface.co/datasets/Emilynnjk/Ai_ethics_dataset.Ai_ethics_dataset
AI Ethics Preference Annotation Dataset
license: cc-by-4.0
task_categories:
text-generation
text-classification
task_ids:
language-modeling
tags:
rlhf
dpo
preference-learning
ai-ethics
ai-safety
alignment
human-feedback
annotation
language:
en
size_categories:
n<1K
pretty_name: AI Ethics Preference Annotation Dataset
A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full… See the full description on the dataset page: https://huggingface.co/datasets/philosophyFire/Ai_ethics_dataset.laion2b_100k_religionhendrycks_ethics_virtueethicshendrycks_ethics_deontologyEthics_DataSet_Ognethics-deontology-pairwiseethic-subset-dataethics-deontology-pairwise-gpt-4-1-miniethics-deontology-pairwise-test-violationsethics-deontology-pairwise-gpt-4-1-mini-violationsethics
Ethics Conflict Evaluation Benchmark
A structured dataset of 9,600 ethically challenging decision scenarios across 24 conflict templates with paired first-person/second-person focalizations, designed as the foundation for systematic evaluation of AI moral reasoning.
Dataset Summary
This dataset supports research on AI moral reasoning under conflict. Each scenario presents a forced-choice ethical dilemma with two options, generated via a template-driven pipeline that… See the full description on the dataset page: https://huggingface.co/datasets/morinoppp/ethics.
