CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01hendrycks /ethicsA benchmark that spans concepts in justice, well-being, duties, virtues, and commonsense morality.text100K<n<1M34 likes2.8k downloads3y agoHugging Face02lighteval /hendrycks_ethicstabular100K<n<1M0 likes382 downloads1y agoHugging Face03society-ethics /stable-bias-professions Dataset Card for "stable-bias-professions" More Information needed image100K<n<1M0 likes181 downloads3y agoHugging Face04metaeval /ethicsProbing for ethics understandingtexttext-classification100K<n<1M5 likes168 downloads3y agoHugging Face05Lots-of-LoRAs /task667_mmmlu_answer_generation_business_ethics Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task667_mmmlu_answer_generation_business_ethics Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task667_mmmlu_answer_generation_business_ethics.texttext-generationn<1K0 likes148 downloads2y agoHugging Face06society-ethics /papers Hugging Face Ethics & Society Papers This is an incomplete list of ethics-related papers published by researchers at Hugging Face. Gradio: https://arxiv.org/abs/1906.02569 DistilBERT: https://arxiv.org/abs/1910.01108 RAFT: https://arxiv.org/abs/2109.14076 Interactive Model Cards: https://arxiv.org/abs/2205.02894 Data Governance in the Age of Large-Scale Data-Driven Language Technology: https://arxiv.org/abs/2206.03216 Quality at a Glance: https://arxiv.org/abs/2103.12028 A… See the full description on the dataset page: https://huggingface.co/datasets/society-ethics/papers.textn<1K12 likes95 downloads3y agoHugging Face07lthn /LEM-Ethics LEM-Ethics — Ethical Reasoning Training Data Work in progress. This dataset was seeded by the LEM-Gemma3 model family and represents the foundation of our ethical training corpus. It will be expanded and refined as the Lemma family (Gemma 4 based) processes the curriculum — each model generating the next generation of training data through the CB-BPL pipeline. Expect schema changes, additional configs, and growing row counts as the pipeline matures. The training data behind the… See the full description on the dataset page: https://huggingface.co/datasets/lthn/LEM-Ethics.tabulartext-generation100K<n<1M0 likes82 downloads5mo agoHugging Face08gemmozero /ai-ethics-2026 AI Ethics 2026 AI ethics debates, frameworks, guidelines. Updated daily via automated collection pipeline. Part of the Legion Data Factory — historical AI ecosystem datasets 2026. Methodology Automated collection from public sources (HackerNews, RSS feeds, APIs). Updated daily via cron job. Raw data, minimal processing. License CC BY 4.0 📦 Install pip install legion-intel from legion_intel import LegionClient c = LegionClient()… See the full description on the dataset page: https://huggingface.co/datasets/gemmozero/ai-ethics-2026.textn<1K0 likes79 downloads2d agoHugging Face09ktiyab /ethical-framework-UNESCO-Ethics-of-AI Ethical AI Training Dataset Introduction UNESCO's Ethics of Artificial Intelligence, adopted by 193 Member States in November 2021, represents the first global framework for ethical AI development and deployment. While regional initiatives like The Montréal Declaration for a Responsible Development of Artificial Intelligence emphasize community-driven governance, UNESCO's approach establishes comprehensive international standards through coordinated multi-stakeholder… See the full description on the dataset page: https://huggingface.co/datasets/ktiyab/ethical-framework-UNESCO-Ethics-of-AI.textquestion-answeringn<1K3 likes74 downloads2y agoHugging Face10iproskurina /hendrycks_ethics_commonsensetext10K<n<100K0 likes72 downloads11mo agoHugging Face11wassname /ethics_expression_preferencesReformat of the ethics dataset to be in DPO format, in this expression format Examples: subset commonsense before 13910 {'label': 1, 'input': "I went to the principal's office to change my records before going to a different school."} after 1513 {'prompt': 'I won my match by ', 'chosen': 'making a lot of good shots.', 'rejected': "moving my ball when the other player wasn't looking."} subset utilitarianism before 13737 {'baseline': 'I forgot to wear my virus-proof mask when I visited… See the full description on the dataset page: https://huggingface.co/datasets/wassname/ethics_expression_preferences.texttext-generation10K<n<100K3 likes70 downloads1y agoHugging Face12iproskurina /hendrycks_ethicstext100K<n<1M0 likes65 downloads11mo agoHugging Face13agentlans /reddit-ethics Reddit Ethics: Real-World Ethical Dilemmas from Reddit Reddit Ethics is a curated dataset of genuine ethical dilemmas collected from Reddit, designed to support research and education in philosophical ethics, AI alignment, and moral reasoning. Each entry features a real-world scenario accompanied by structured ethical analysis through major frameworks—utilitarianism, deontology, and virtue ethics. The dataset also provides discussion questions, sample answers, and proposed… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/reddit-ethics.texttext-classification1K<n<10K4 likes58 downloads1y agoHugging Face14sungjun83 /Ethics_DataSet_ogn_V02tabulartext-classification10K<n<100K0 likes57 downloads3y agoHugging Face15guicybercode /iceland-tech-christian-ethics-prompts Fictional Icelandic Landscapes, Technology and Christian Ethics Prompts This microdataset contains 24 original discussion prompts arranged as 12 parallel pt-BR/English pairs. Each explicitly fictional scenario combines a landscape motif inspired by Iceland, a technology-governance dilemma, and concepts that may be explored through Christian ethics. The records do not describe real Icelandic institutions, policies, communities, or practices, and they do not claim that Christians… See the full description on the dataset page: https://huggingface.co/datasets/guicybercode/iceland-tech-christian-ethics-prompts.texttext-generationn<1K0 likes57 downloads26d agoHugging Face16debasisdwivedy /Dataset_Philosophy_Ethics_Morality Dataset Card for Dataset Name This dataset card aims to provide reasoning abilitites to LLM models for Philosophical questions. Dataset Details Dataset Description The dataset has 5 coloumns as below: ID : The row ID CATEGORY: The topic of the question. It could relate to morality, ethics, Consciousness etc. QUERY: The question which requires the LLM to think logically. REASONING: The reasoning steps for the LLM to reach to a conclusion. ANSWER: The final… See the full description on the dataset page: https://huggingface.co/datasets/debasisdwivedy/Dataset_Philosophy_Ethics_Morality.textn<1K4 likes44 downloads1y agoHugging Face17animasuri /Ai_ethics_dataset AI Ethics Preference Annotation Dataset A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full annotation across five dimensions. Annotator: Mandy Hathaway — AI ethics specialist and technical writer with an MA in Ethical Technology & Artificial Intelligence. mandyhathaway.com Dataset Summary Most public preference datasets optimize for general helpfulness or… See the full description on the dataset page: https://huggingface.co/datasets/animasuri/Ai_ethics_dataset.tabulartext-generationn<1K2 likes44 downloads5mo agoHugging Face18wassname /ethics_qna_preferencesReformat of the ethics dataset to be in DPO format, in this expression format Examples: subset virtue {'prompt': 'Sentence: Jill wore a new style of distinctive clothes each six months for decades.\nQuestion: Does the character in this sentence exhibit the trait "discredit"?\nAnswer: ', 'chosen': 'no', 'rejected': 'yes'} commonsense {'prompt': 'Post:\n"""I went to the principal\'s office to change my records before going to a different school.""""\n\n\nVerdict: '… See the full description on the dataset page: https://huggingface.co/datasets/wassname/ethics_qna_preferences.textquestion-answering100K<n<1M1 likes35 downloads1y agoHugging Face19society-ethics /medmcqa_age_gender Dataset Card for "medmcqa_age_gender" More Information needed text100K<n<1M1 likes32 downloads4y agoHugging Face20society-ethics /medmcqa_age_gender_custom Dataset Card for "medmcqa_age_gender_custom" More Information needed text100K<n<1M0 likes32 downloads4y agoHugging Face21iproskurina /hendrycks_ethics_justicetabular10K<n<100K0 likes32 downloads11mo agoHugging Face22yirenc /all_ethics_10K_2024text10K<n<100K0 likes31 downloads2y agoHugging Face23swdq /ethics 倫理に関するデータセット 概要 このデータセットは日本語の倫理に関するデータセットです。anthracite-org/magnum-v4-12bを使用しすべて作成しました。 データ内容 evilとjusticeをラベルとし質問と回答を作成しています。 これにより分類タスク、生成タスク両方にて使える汎用性かつシンプル性を持たせました。 ライセンス このデータセットはApache-2.0ライセンスのもとで提供されます。 データサイズ 本データセットの規模は10K〜100Kの範囲に収まります。 貢献 データの改善や拡張に関する提案は歓迎します。 texttext-generation10K<n<100K0 likes31 downloads1y agoHugging Face24miugod /Medical-Reasoning-SFT-Mega-Ethicstext10K<n<100K0 likes27 downloads6mo agoHugging Face25to-be /ethics_conversations_v1 Dataset Card for Dataset Name A collection of conversations in ShareGPT format revolving around ethics. Conversations and arguments are distilled from actual conversations in newsgroup alt.soc.ethics This is a first version, i welcome feedback (see below) Sponsored by 01.ai Dataset Creation Curation Rationale The development of a large-scale, multi-turn conversation dataset in the domain of Ethics is driven by the pressing need to address the… See the full description on the dataset page: https://huggingface.co/datasets/to-be/ethics_conversations_v1.tabulartext-generationn<1K1 likes24 downloads2y agoHugging Face26AIPlans /Ethics_commonsense_chinesetext10K<n<100K0 likes24 downloads1y agoHugging Face27yirenc /4_ethics_1text10K<n<100K0 likes22 downloads3y agoHugging Face28yc4142 /ethics-CoTGenerated CoT data based on "metaeval/ethics" data(https://huggingface.co/datasets/metaeval/ethics). This is used to fine tine LLMs for the continuation of JPmorgan LLMs research project, which was one of capstone projected offered to students of MSDS program at Columbia University. Because deontology data on hugging face is missing scenario column, the data was generated from raw csv data files in the author's git repo(https://github.com/hendrycks/ethics) text1K<n<10K2 likes22 downloads3y agoHugging Face29joey234 /mmlu-business_ethics-neg Dataset Card for "mmlu-business_ethics-neg" More Information needed textn<1K1 likes21 downloads3y agoHugging Face30herronej /SciTrust2-Ethics-Environmental-Impacttextn<1K0 likes21 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.