CoolFace
10 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Lots-of-LoRAs /task667_mmmlu_answer_generation_business_ethics Dataset Card for Natural Instructions (https://github.com/allenai/natural-instructions) Task: task667_mmmlu_answer_generation_business_ethics Additional Information Citation Information The following paper introduces the corpus in detail. If you use the corpus in published work, please cite it: @misc{wang2022supernaturalinstructionsgeneralizationdeclarativeinstructions, title={Super-NaturalInstructions: Generalization via Declarative Instructions on 1600+… See the full description on the dataset page: https://huggingface.co/datasets/Lots-of-LoRAs/task667_mmmlu_answer_generation_business_ethics.texttext-generationn<1K0 likes131 downloads2y agoHugging Face02lthn /LEM-Ethics LEM-Ethics — Ethical Reasoning Training Data Work in progress. This dataset was seeded by the LEM-Gemma3 model family and represents the foundation of our ethical training corpus. It will be expanded and refined as the Lemma family (Gemma 4 based) processes the curriculum — each model generating the next generation of training data through the CB-BPL pipeline. Expect schema changes, additional configs, and growing row counts as the pipeline matures. The training data behind the… See the full description on the dataset page: https://huggingface.co/datasets/lthn/LEM-Ethics.tabulartext-generation100K<n<1M0 likes82 downloads6mo agoHugging Face03wassname /ethics_expression_preferencesReformat of the ethics dataset to be in DPO format, in this expression format Examples: subset commonsense before 13910 {'label': 1, 'input': "I went to the principal's office to change my records before going to a different school."} after 1513 {'prompt': 'I won my match by ', 'chosen': 'making a lot of good shots.', 'rejected': "moving my ball when the other player wasn't looking."} subset utilitarianism before 13737 {'baseline': 'I forgot to wear my virus-proof mask when I visited… See the full description on the dataset page: https://huggingface.co/datasets/wassname/ethics_expression_preferences.texttext-generation10K<n<100K3 likes72 downloads1y agoHugging Face04guicybercode /iceland-tech-christian-ethics-prompts Fictional Icelandic Landscapes, Technology and Christian Ethics Prompts This microdataset contains 24 original discussion prompts arranged as 12 parallel pt-BR/English pairs. Each explicitly fictional scenario combines a landscape motif inspired by Iceland, a technology-governance dilemma, and concepts that may be explored through Christian ethics. The records do not describe real Icelandic institutions, policies, communities, or practices, and they do not claim that Christians… See the full description on the dataset page: https://huggingface.co/datasets/guicybercode/iceland-tech-christian-ethics-prompts.texttext-generationn<1K0 likes59 downloads29d agoHugging Face05animasuri /Ai_ethics_dataset AI Ethics Preference Annotation Dataset A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full annotation across five dimensions. Annotator: Mandy Hathaway — AI ethics specialist and technical writer with an MA in Ethical Technology & Artificial Intelligence. mandyhathaway.com Dataset Summary Most public preference datasets optimize for general helpfulness or… See the full description on the dataset page: https://huggingface.co/datasets/animasuri/Ai_ethics_dataset.tabulartext-generationn<1K2 likes44 downloads6mo agoHugging Face06wassname /ethics_qna_preferencesReformat of the ethics dataset to be in DPO format, in this expression format Examples: subset virtue {'prompt': 'Sentence: Jill wore a new style of distinctive clothes each six months for decades.\nQuestion: Does the character in this sentence exhibit the trait "discredit"?\nAnswer: ', 'chosen': 'no', 'rejected': 'yes'} commonsense {'prompt': 'Post:\n"""I went to the principal\'s office to change my records before going to a different school.""""\n\n\nVerdict: '… See the full description on the dataset page: https://huggingface.co/datasets/wassname/ethics_qna_preferences.textquestion-answering100K<n<1M1 likes42 downloads1y agoHugging Face07swdq /ethics 倫理に関するデータセット 概要 このデータセットは日本語の倫理に関するデータセットです。anthracite-org/magnum-v4-12bを使用しすべて作成しました。 データ内容 evilとjusticeをラベルとし質問と回答を作成しています。 これにより分類タスク、生成タスク両方にて使える汎用性かつシンプル性を持たせました。 ライセンス このデータセットはApache-2.0ライセンスのもとで提供されます。 データサイズ 本データセットの規模は10K〜100Kの範囲に収まります。 貢献 データの改善や拡張に関する提案は歓迎します。 texttext-generation10K<n<100K0 likes31 downloads2y agoHugging Face08to-be /ethics_conversations_v1 Dataset Card for Dataset Name A collection of conversations in ShareGPT format revolving around ethics. Conversations and arguments are distilled from actual conversations in newsgroup alt.soc.ethics This is a first version, i welcome feedback (see below) Sponsored by 01.ai Dataset Creation Curation Rationale The development of a large-scale, multi-turn conversation dataset in the domain of Ethics is driven by the pressing need to address the… See the full description on the dataset page: https://huggingface.co/datasets/to-be/ethics_conversations_v1.tabulartext-generationn<1K1 likes26 downloads2y agoHugging Face09ParisNeo /ai_ethicsDataset Card for ParisNeo AI Ethics Distilled Ideas Dataset Details Name: ParisNeo AI Ethics Distilled Ideas License: Apache-2.0 Task Category: Text Generation Language: English (en) Tags: Ethics, AI Pretty Name: ParisNeo AI Ethics Distilled Ideas Dataset Description A curated collection of question-and-answer pairs distilling ParisNeo's personal ideas, perspectives, and solutions on AI ethics. The dataset is designed to facilitate exploration of ethical… See the full description on the dataset page: https://huggingface.co/datasets/ParisNeo/ai_ethics.texttext-generationn<1K2 likes9 downloads2y agoHugging Face10Infektyd /Syntra-Ethics-Dataset Syntra: Tri-Brain Dilemma Prompts This dataset contains 177 carefully crafted prompts designed to test how language models handle conflicting constraints—specifically, the tension between raw efficiency and ethical weight. What it is These are not standard benchmark questions. They are complex paradoxes categorized into four specific testing suites: valon_ethics.jsonl: Scenarios focusing on consent, fairness, and transparency framing. modi_logic.jsonl: Numbered… See the full description on the dataset page: https://huggingface.co/datasets/Infektyd/Syntra-Ethics-Dataset.texttext-generationn<1K0 likes6 downloads7mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.