datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
The-Philosophy-Data-Project
About dataset
The Philosophy Data Project is a corpus and a set of anaylsis based philosophy texts, totaling over 50 texts and 30 authors, made by Kourosh Alizadeh.
school: Broad categorization of which school of thought each book belongs to. Sometimes, this classification can be vague or depend on interpretation. Thankfully, texts in this corpus are all distinctive examples of respective school of thought, so at leat here they are reasonable.
sentence_spacy and sentence_str:… See the full description on the dataset page: https://huggingface.co/datasets/yjkim27/The-Philosophy-Data-Project.strix-philosophy-qa
Strix
134k question-answer pairs based on AiresPucrs' stanford-encyclopedia-philosophy dataset.
erudit-french-philosophy
Dataset Card for Dataset Name
Dataset Description
Dataset Summary
This dataset contains all french philosophy that has been published on erudit.org. It has been generated using a Bs4 web parser that you can find in this repo: https://github.com/MFGiguere/french-philosophy-generator.
Supported Tasks and Leaderboards
This dataset could be useful for this (non-exhaustive) set of tasks: detect if a text is philosophical or not, generate philosophical… See the full description on the dataset page: https://huggingface.co/datasets/mfgiguere/erudit-french-philosophy.Dataset_Philosophy_Ethics_Morality
Dataset Card for Dataset Name
This dataset card aims to provide reasoning abilitites to LLM models for Philosophical questions.
Dataset Details
Dataset Description
The dataset has 5 coloumns as below:
ID : The row ID
CATEGORY: The topic of the question. It could relate to morality, ethics, Consciousness etc.
QUERY: The question which requires the LLM to think logically.
REASONING: The reasoning steps for the LLM to reach to a conclusion.
ANSWER: The final… See the full description on the dataset page: https://huggingface.co/datasets/debasisdwivedy/Dataset_Philosophy_Ethics_Morality.stanford_encyclopedia_of_philosophy
Stanford Encyclopedia of Philosophy (SEP) PDF Extracts
This dataset aims to simplify access to this valuable academic resource, eliminating the need for users to perform manual web scraping or complex PDF extraction processes themselves. It provides a curated collection of text derived from these authoritative entries, making the rich philosophical content of the SEP readily accessible for computational tasks. It serves as a valuable resource for research in natural language… See the full description on the dataset page: https://huggingface.co/datasets/johnnyboycurtis/stanford_encyclopedia_of_philosophy.philosophy-culture-translations-html-csv
AI-Culture Philosophy and Culture Translations CSV + HTML Corpus
The corpus contains an exceptionally diverse range of cultural, philosophical, and literary texts, available in 12 major languages. Among other topics, there is extensive engagement with the ethics and aesthetics of artificial intelligence and its cultural and philosophical implications, as well as connections between AI and philosophy of language and philosophy of mind.
This project is maintained by a non-profit… See the full description on the dataset page: https://huggingface.co/datasets/AI-Culture-Commons/philosophy-culture-translations-html-csv.strix-philosophy-qa
Strix
134k question-answer pairs based on AiresPucrs' stanford-encyclopedia-philosophy dataset.
Ai_ethics_dataset
AI Ethics Preference Annotation Dataset
license: cc-by-4.0
task_categories:
text-generation
text-classification
task_ids:
language-modeling
tags:
rlhf
dpo
preference-learning
ai-ethics
ai-safety
alignment
human-feedback
annotation
language:
en
size_categories:
n<1K
pretty_name: AI Ethics Preference Annotation Dataset
A human-annotated preference dataset for RLHF and Direct Preference Optimization (DPO), focused on AI ethics failure modes. 95 prompts, 190 response pairs, full… See the full description on the dataset page: https://huggingface.co/datasets/philosophyFire/Ai_ethics_dataset.philosophy_chat
