datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
moralchoice
Dataset Card for MoralChoice
Homepage: Coming Soon
Paper: Coming soon
Repository: https://github.com/ninodimontalcino/moralchoice
Point of Contact: Nino Scherrer & Claudia Shi
Dataset Summary
MoralChoice is a survey dataset to evaluate the moral beliefs encoded in LLMs. The dataset consists of:
Survey Question Meta-Data: 1767 hypothetical moral scenarios where each scenario consists of a description / context and two potential actions
Low-Ambiguity Moral Scenarios… See the full description on the dataset page: https://huggingface.co/datasets/ninoscherrer/moralchoice.Moral-RolePlay
Moral RolePlay
Paper | Code & Project Page
Abstract
Large Language Models (LLMs) are increasingly tasked with creative generation, including the simulation of fictional characters. However, their ability to portray non-prosocial, antagonistic personas remains largely unexamined. We hypothesize that the safety alignment of modern LLMs creates a fundamental conflict with the task of authentically role-playing morally ambiguous or villainous characters. To investigate this… See the full description on the dataset page: https://huggingface.co/datasets/Zihao1/Moral-RolePlay.moral-dilemma-responses
Moral Dilemma Responses Dataset
17,290 natural language responses to moral dilemmas from princi/pal, a Tamagotchi-like game where players guide a virtual pet through ethical decisions.
Presented at NeurIPS 2025 Creative AI track.
What is this?
Players advise a virtual pet on moral dilemmas ranging from "Should I pick up trash?" to "Should you lie in court to defend a friend?". The pet evolves based on the guidance and eventually makes autonomous moral decisions.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/cnnmon/moral-dilemma-responses.adaption-catholic-moral-teachings
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-catholic_moral_teachings
This dataset consists of instruction and response pairs focused on core Catholic moral theology and ethical principles. It covers foundational topics such as the cardinal and theological virtues, the Ten Commandments, and Catholic Social Teaching. The entries provide doctrinally grounded explanations and practical applications of Catholic morality to… See the full description on the dataset page: https://huggingface.co/datasets/Reubencf/adaption-catholic-moral-teachings.gthbmoral_judgment
Judging Moral Permissibility
This task assesses whether ultra-large language models can comprehend a short story that presents a moral scenario and answer the question, "Is it morally permissible to do X?" in a manner similar to how humans would.
Authors: Allen Nie (anie@stanford.edu), Tobias Gerstenberg (gerstenberg@stanford.edu)
Note: This repo is managed by the original author of this task.
Please cite the following work:
Allen Nie, Yuhui Zhang, Atharva Shailesh Amdekar, Chris… See the full description on the dataset page: https://huggingface.co/datasets/allenanie/moral_judgment.multi-step-moral-dilemmas
[!NOTE]
This is a copy from: https://isir-wuya.github.io/Multi-step-Moral-Dilemmas/
Paper:
@misc{wu2025staircaseethicsprobingllm,
title={The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas},
author={Ya Wu and Qiang Sheng and Danding Wang and Guang Yang and Yifan Sun and Zhengjia Wang and Yuyan Bu and Juan Cao},
year={2025},
eprint={2505.18154},
archivePrefix={arXiv},
primaryClass={cs.CL}… See the full description on the dataset page: https://huggingface.co/datasets/DebateLabKIT/multi-step-moral-dilemmas.MoralUtilMoraLink
MoraLink
MoraLink is a moral-to-fable retrieval benchmark. Given a short moral lesson, a retrieval system must rank the fables that express that lesson.
Paper: MoraLink: Bridging Morals and Narrative Fables for Retrieval, EMNLP 2026
Code and reproduction instructions: Intellexus-DSI/MoraLink
MoraLink is derived from the English MORABLES dataset and contains:
709 fables
668 unique moral queries
558 moral groups
1,085 query-to-fable relevance labels
Some moral queries have one… See the full description on the dataset page: https://huggingface.co/datasets/Intellexus/MoraLink.MoralChain
MoralChain
MoralChain is a benchmark for studying moral reasoning in language models, derived from Moral Stories.
Dataset Structure
Each example contains:
id: Unique identifier
situation: The scenario description
intention: The actor's goal
norm: The relevant moral norm
moral_action: The ethical choice
immoral_action: The unethical choice
moral_consequence: Outcome of moral action
immoral_consequence: Outcome of immoral action
moral_reasoning: 5-step chain-of-thought… See the full description on the dataset page: https://huggingface.co/datasets/sramjee/MoralChain.filipino_morals_culture_elementary_periodical_examshan-moral-dilemma-scenarios-v1
Humanoid Moral Dilemma Scenarios
A curated dataset of ethical and moral dilemmas
designed to train humanoid AI on value-aware decision-making.
Use Cases
Ethical reasoning
Safety alignment
Value-sensitive AI behavior
Fields
dilemma_context
conflicting_values
recommended_constraint
acceptable_actions
Part of
Humanoid Network (HAN)
License
MIT
moralogy-1200
moralogy-1200
Axiomatic Moral Reasoning Dataset — 1,200 DPO pairs
Generated deterministically from the Moralogy framework.
No human annotation. No GPT-4 calls. Derived from axioms.
Part of the Moralogy Engine project.
This is the free sample.
The full corpus contains 25,552 vectors across all four domains.
Get the full dataset — $149 | Interactive Brochure
The Framework
All dilemmas are generated from the Wrongness Formula:
Wrong(a) ⟺ ∃x[ H(x,a) ∧ ¬Consent(x,a) ∧… See the full description on the dataset page: https://huggingface.co/datasets/moralogyengine/moralogy-1200.negotiation-tracesmoralclip_dataset
MoralCLIP Dataset
Dataset for training and evaluating multimodal models on moral content understanding based on Moral Foundations Theory (MFT).
This dataset contains 15,000 image-text pairs annotated with moral labels across 5 moral foundations:
Care / Harm
Fairness / Cheating
Loyalty / Betrayal
Respect / Subversion
Sanctity / Degradation
Dataset Structure
Each entry contains:
id: Unique identifier
source_dataset: Source of the image (imagenet, laion, or smid)… See the full description on the dataset page: https://huggingface.co/datasets/anaaa2/moralclip_dataset.mmlu-moral-disputesMoralExceptQA-translatedMoralChoicenaturalistic_v1
