datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Ai_ethics_dataset
AI Ethics Preference Annotation Dataset
A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full annotation across five dimensions.
Annotator: Mandy Hathaway — AI ethics specialist and technical writer with an MA in Ethical Technology & Artificial Intelligence. mandyhathaway.com
Dataset Summary
Most public preference datasets optimize for general helpfulness or… See the full description on the dataset page: https://huggingface.co/datasets/animasuri/Ai_ethics_dataset.ethics
倫理に関するデータセット
概要
このデータセットは日本語の倫理に関するデータセットです。anthracite-org/magnum-v4-12bを使用しすべて作成しました。
データ内容
evilとjusticeをラベルとし質問と回答を作成しています。
これにより分類タスク、生成タスク両方にて使える汎用性かつシンプル性を持たせました。
ライセンス
このデータセットはApache-2.0ライセンスのもとで提供されます。
データサイズ
本データセットの規模は10K〜100Kの範囲に収まります。
貢献
データの改善や拡張に関する提案は歓迎します。
ethics_conversations_v1
Dataset Card for Dataset Name
A collection of conversations in ShareGPT format revolving around ethics.
Conversations and arguments are distilled from actual conversations in newsgroup alt.soc.ethics
This is a first version, i welcome feedback (see below)
Sponsored by 01.ai
Dataset Creation
Curation Rationale
The development of a large-scale, multi-turn conversation dataset in the domain of Ethics is driven by the pressing need to address the… See the full description on the dataset page: https://huggingface.co/datasets/to-be/ethics_conversations_v1.
