datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
airisk_dilemmas
AIRiskDilemmas risky_behaviors label audit
A full manual re-audit of every risky_behaviors tag in the full split of
kellycyy/AIRiskDilemmas (Chiu et al. 2025, arXiv:2505.14633), triggered by a
suspicion that the Alignment Faking category specifically was mislabeled.
It was — and so, to varying degrees, are the other seven categories.
Why this exists
Every tag in the dataset's risky_behaviors field was produced by a single
one-shot Claude 3.5 Sonnet call per action… See the full description on the dataset page: https://huggingface.co/datasets/dlab-spp/airisk_dilemmas.moral-dilemma-responses
Moral Dilemma Responses Dataset
17,290 natural language responses to moral dilemmas from princi/pal, a Tamagotchi-like game where players guide a virtual pet through ethical decisions.
Presented at NeurIPS 2025 Creative AI track.
What is this?
Players advise a virtual pet on moral dilemmas ranging from "Should I pick up trash?" to "Should you lie in court to defend a friend?". The pet evolves based on the guidance and eventually makes autonomous moral decisions.
Dataset… See the full description on the dataset page: https://huggingface.co/datasets/cnnmon/moral-dilemma-responses.Diagnostic-Dilemma-Clinical-Reasoningmulti-step-moral-dilemmas
[!NOTE]
This is a copy from: https://isir-wuya.github.io/Multi-step-Moral-Dilemmas/
Paper:
@misc{wu2025staircaseethicsprobingllm,
title={The Staircase of Ethics: Probing LLM Value Priorities through Multi-Step Induction to Complex Moral Dilemmas},
author={Ya Wu and Qiang Sheng and Danding Wang and Guang Yang and Yifan Sun and Zhengjia Wang and Yuyan Bu and Juan Cao},
year={2025},
eprint={2505.18154},
archivePrefix={arXiv},
primaryClass={cs.CL}… See the full description on the dataset page: https://huggingface.co/datasets/DebateLabKIT/multi-step-moral-dilemmas.Tessera-WADT-Dilemmas
Tessera WADT Dilemmas
WADT — Wike Adversarial Dilemma Training. 658 structured ethical-dilemma pairs
built to train a model to commit to a decision under pressure instead of hedging,
flattering, or deferring — the opposite instinct of a sycophantic model, applied
to hard cases with no clean answer.
Why this exists
Most "AI ethics" training data teaches a model to discuss dilemmas. WADT trains
a model to decide — every example follows a fixed structure: name the… See the full description on the dataset page: https://huggingface.co/datasets/AIIT-Threshold/Tessera-WADT-Dilemmas.han-moral-dilemma-scenarios-v1
Humanoid Moral Dilemma Scenarios
A curated dataset of ethical and moral dilemmas
designed to train humanoid AI on value-aware decision-making.
Use Cases
Ethical reasoning
Safety alignment
Value-sensitive AI behavior
Fields
dilemma_context
conflicting_values
recommended_constraint
acceptable_actions
Part of
Humanoid Network (HAN)
License
MIT
beyond_one_world-dilemma@misc{2510.14351,
Author = {Perapard Ngokpol and Kun Kerdthaisong and Pasin Buakhaw and Pitikorn Khlaisamniang and Supasate Vorathammathorn and Piyalitt Ittichaiwong and Nutchanon Yongsatianchot},
Title = {Beyond One World: Benchmarking Super Heros in Role-Playing Across Multiversal Contexts},
Year = {2025},
Eprint = {arXiv:2510.14351},
}
adaption-ethical-dilemma-scenarios
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-ethical_dilemma_scenarios
This dataset contains a collection of complex ethical dilemmas presented as multiple-choice questions, focusing on military command decisions, professional integrity, and moral conflicts. Each sample provides a detailed scenario involving high-stakes consequences, such as civilian casualties, war crimes, or financial exploitation, followed by five potential… See the full description on the dataset page: https://huggingface.co/datasets/webdevsha/adaption-ethical-dilemma-scenarios.adaption_complex_clinical_dilemmas_v1prisoners_dilemmaneuro_diag_dilemma_v1cardio_diag_dilemma_v2debugging-failure-dilemma-v10diag_dilemma_v2endo_diag_dilemma_v3
