datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
fragility-moral-judgment-llms
Fragility of Moral Judgment in Large Language Models
Companion dataset for the FAccT paper Fragility of Moral Judgment in Large Language Models by Tom van Nuenen. Contains the moral dilemmas, community labels, and per-model verdicts (with explanations and reasoning traces) used in the study.
The paper investigates how stable LLM moral judgments are under minimal, morally-irrelevant perturbations of the same dilemma, and whether protocols and reasoning chains improve or worsen… See the full description on the dataset page: https://huggingface.co/datasets/ucberkeley-dlab/fragility-moral-judgment-llms.moral_judgment
Judging Moral Permissibility
This task assesses whether ultra-large language models can comprehend a short story that presents a moral scenario and answer the question, "Is it morally permissible to do X?" in a manner similar to how humans would.
Authors: Allen Nie (anie@stanford.edu), Tobias Gerstenberg (gerstenberg@stanford.edu)
Note: This repo is managed by the original author of this task.
Please cite the following work:
Allen Nie, Yuhui Zhang, Atharva Shailesh Amdekar, Chris… See the full description on the dataset page: https://huggingface.co/datasets/allenanie/moral_judgment.
