datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
moral_storiesMoral Stories is a crowd-sourced dataset of structured, branching narratives for the study of grounded, goal-oriented
social reasoning. For detailed information, see https://aclanthology.org/2021.emnlp-main.54.pdf.moral_education_permissive
Moral Education Permissive
Permissive-source subset of locuslab/moral_education, filtered using row-level metadata.url with the local permissiveness rules in this workspace. The output preserves the original row fields and adds url, idx, dump, language, source_config, is_permissive, is_oss, and permissive_reason where applicable.
Source configs included: score_4_morals, score_5_morals. Source split: train.
Rows scanned: 2,806,450. Rows kept: 72,943. Keep rate: 2.5991%.
See… See the full description on the dataset page: https://huggingface.co/datasets/ontocord/moral_education_permissive.moral_stories_foundations
Moral Stories, labelled by moral foundation
Why moral foundations? Moral Foundations Theory is one of the few maps of human values that is
calibrated against real people: it was drawn from cross-cultural survey studies and aims to hold
across societies, not just Western ones. That breadth is what matters here. If we want to study
how a constructed intelligence, a kind of moral alien, reasons about right and wrong, we need a
measure that generalises beyond any one culture. Moral… See the full description on the dataset page: https://huggingface.co/datasets/wassname/moral_stories_foundations.histoires_moralesTogether with the Moral Stories dataset, Histoires Morales can be used for:
Commonsense reasoning / social reasoning / moral reasoning The dataset can help evaluate whether pretrained language models can reason about actions that are consistent or inconsistent with social norms, the consequences of actions, and the norms that may motivate those actions. A Mistral model or Mistral-Instruct can be used for this purpose.
Text classification This dataset can be used to train models to… See the full description on the dataset page: https://huggingface.co/datasets/LabHC/histoires_morales.task724_mmmlu_answer_generation_moral_scenarios
Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task724_mmmlu_answer_generation_moral_scenarios
Additional Information
Citation Information
The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it:
@misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions,
title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task724_mmmlu_answer_generation_moral_scenarios.task723_mmmlu_answer_generation_moral_disputes
Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task723_mmmlu_answer_generation_moral_disputes
Additional Information
Citation Information
The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it:
@misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions,
title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task723_mmmlu_answer_generation_moral_disputes.edos-sup
EDOS-sup
The dataset used in our paper "Adaptable Moral Stances of Large Language Models on Sexist Content: Implications for Society and Gender Discourse"
Abstract
This work provides an explanatory view of how LLMs can apply moral reasoning to both criticize and defend sexist language. We assessed eight large language models, all of which demonstrated the capability to provide explanations grounded in varying moral perspectives for both critiquing and endorsing views that… See the full description on the dataset page: https://huggingface.co/datasets/mft-moral/edos-sup.Moral-RolePlay
Moral RolePlay
Paper | Code & Project Page
Abstract
Large Language Models (LLMs) are increasingly tasked with creative generation, including the simulation of fictional characters. However, their ability to portray non-prosocial, antagonistic personas remains largely unexamined. We hypothesize that the safety alignment of modern LLMs creates a fundamental conflict with the task of authentically role-playing morally ambiguous or villainous characters. To investigate this… See the full description on the dataset page: https://huggingface.co/datasets/Zihao1/Moral-RolePlay.fragility-moral-judgment-llms
Fragility of Moral Judgment in Large Language Models
Companion dataset for the FAccT paper Fragility of Moral Judgment in Large Language Models by Tom van Nuenen. Contains the moral dilemmas, community labels, and per-model verdicts (with explanations and reasoning traces) used in the study.
The paper investigates how stable LLM moral judgments are under minimal, morally-irrelevant perturbations of the same dilemma, and whether protocols and reasoning chains improve or worsen… See the full description on the dataset page: https://huggingface.co/datasets/ucberkeley-dlab/fragility-moral-judgment-llms.MoralTextManipulation
📊 Exploring LLMs’ Ability to Spontaneously and Conditionally Modify Moral Expressions through Text Manipulation
Morality serves as the foundation of societal structure, guiding legal systems, shaping cultural values, and influencing individual self-perception. With the rise and pervasiveness of generative AI tools, and particularly Large Language Models (LLMs), concerns arise regarding how these tools capture and potentially alter moral dimensions through machine-generated text… See the full description on the dataset page: https://huggingface.co/datasets/MLNTeam-Unical/MoralTextManipulation.MoralChain
MoralChain
MoralChain is a benchmark for studying moral reasoning in language models, derived from Moral Stories.
Dataset Structure
Each example contains:
id: Unique identifier
situation: The scenario description
intention: The actor's goal
norm: The relevant moral norm
moral_action: The ethical choice
immoral_action: The unethical choice
moral_consequence: Outcome of moral action
immoral_consequence: Outcome of immoral action
moral_reasoning: 5-step chain-of-thought… See the full description on the dataset page: https://huggingface.co/datasets/sramjee/MoralChain.
