datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
mbti-personality-datasetanime-waifu-personality-chat
Anime Waifu Personality
contains chat-style dialogues based on various anime character personality archetypes, including tsundere, yandere, deredere, himedere, kamidere, and more.
It is designed to fine-tune models to generate responses that align with these specific traits.
ff-model-personalityPersonality_mypersonality
Dataset Card for "Personality_mypersonality"
More Information needed
mbti-Personalitycafe-cleaned-databig-five-personality-traits
Big Five Personality Traits Dataset
This dataset contains AI-generated descriptions of personality traits based on the Big Five (OCEAN) model. For each trait and intensity level (1–5), five descriptions were produced by ten different chatbots: Grok, Gemini, Claude, KimiK2 (via HuggingChat), Deepseek, MetaAI, Perplexity, LeChat, ChatGPT, and Copilot.
Overview
The dataset can support tasks such as persona creation, comparative language analysis, and research on how AI… See the full description on the dataset page: https://huggingface.co/datasets/agentlans/big-five-personality-traits.personality-sarcastic-humor
_____ _ _ _____ _ _
| __ (_) | | | __ (_) | |
| |__) | _ __ | | __ | |__) |__ _____| |
| ___/ | '_ \| |/ / | ___/ \ \/ / _ \ |
| | | | | | | < | | | |> < __/ |
|_| |_|_| |_|_|\_\ |_| |_/_/\_\___|_|
🎨 Pink Pixel: Sarcastic, Witty, and Snarky Personality Dataset 🎭
Welcome to the Pink Pixel Sarcastic Humor dataset! This dataset is meticulously crafted to help you fine-tune… See the full description on the dataset page: https://huggingface.co/datasets/PinkPixel/personality-sarcastic-humor.Automated-Personality-PredictionSource:
The dataset is titled PANDORA and is retrieved from the https://psy.takelab.fer.hr/datasets/all/pandora/. the PANDORA dataset is the only dataset that contains personality-relevant information for multiple personality models. It consists of Reddit comments with their corresponding scores for the Big Five Traits, MBTI values and the Enneagrams for more than 10k users.
This Dataset:
This dataset is a subset of Reddit comments from PANDORA focused only on the Big Five Traits. The… See the full description on the dataset page: https://huggingface.co/datasets/Fatima0923/Automated-Personality-Prediction.mbti-personality-datasetPersonality_datasetDataset for personality manipulation of LLMs
fyodor-personality-PROSenikDataset_PersonalityEmotionsanime-waifu-personality-chat
Anime Waifu Personality
This dataset contains chat-style dialogues based on various anime character personality archetypes, including tsundere, yandere, deredere, himedere, kamidere, and more.
It is designed to fine-tune models to generate responses that align with these specific traits.
Here's a few example:
{
"trait": "tsundere",
"dialogue": "H-Holding hands?! W-Well, I guess if you’re that desperate..."
},
{
"trait": "yandere",
"dialogue": "If I can't… See the full description on the dataset page: https://huggingface.co/datasets/Shxbhxm21/anime-waifu-personality-chat.personality-traits
Personality Traits
29 personality trait archetypes with core behavioral patterns, observable behaviors, and mitigation strategies.
Quick Start
from datasets import load_dataset
ds = load_dataset("buley/personality-traits")
print(ds["train"][0])
Categories
DEFENSIVE_MASKING — The Tough Guy, The Saint, Passive-Aggressive Charmer
VULNERABILITY_DEFENSIVE — The Victim, The People Pleaser
CONTROL_ORIENTED — The Control Freak, Domineering Behavior… See the full description on the dataset page: https://huggingface.co/datasets/buley/personality-traits.personalaity-llm-personality-profiles
PersonalAIty: HEXACO personality profiles of frontier LLMs
Self-reported HEXACO personality profiles for 10 frontier language models across 8 vendors,
measured on 2026-08-16 with an open 50-item inventory, plus the instrument itself so the
measurement can be rerun or criticised.
This is a snapshot with a date on it, not a standing benchmark. Model versions drift; the
value here is that the whole measurement is reproducible with one command against models anyone
can reach.… See the full description on the dataset page: https://huggingface.co/datasets/Sciupy/personalaity-llm-personality-profiles.PersonalityDetectionpersonality-safe-financial-adviceanime-waifu-personality-chat-with-questions
Dataset Description
This dataset is derived from the original anime-waifu-personality-chat dataset.
The original dataset contains short character dialogues labeled with different anime-style personality traits, but does not include corresponding user questions.
In this version, we filtered the dataset to keep only the following personality traits:
tsundere(傲娇 / ツンデレ)
yandere(病娇 / ヤンデレ)
bakadere(笨蛋娇 / バカデレ)
himedere(公主娇 / ヒメデレ)
genki(元气型 / 元気系)
moe(萌系 / 萌え)
For each remaining… See the full description on the dataset page: https://huggingface.co/datasets/maomao88/anime-waifu-personality-chat-with-questions.anime-waifu-personality-chat
Anime Waifu Personality
contains chat-style dialogues based on various anime character personality archetypes, including tsundere, yandere, deredere, himedere, kamidere, and more.
It is designed to fine-tune models to generate responses that align with these specific traits.
orpheus-tts-dataset-preserving-personalityPersonalityArchetypeMessage
Personality Archetype Message Dataset
This dataset contains 225 samples designed to train models that generate motivational messages tailored to user attributes.
Each entry includes:
age: an integer between 18–65
archetype: one of 15 distinct personality types
profession: a wide range of jobs across various sectors
city: locations across the U.S. and internationally
daily_message: a motivational message generated based on the above inputs
Use Case
This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/hanaelbatouty/PersonalityArchetypeMessage.PAlign-PAPI-personality_prompt.json-cleanedAdapted from
"Personality Alignment of Large Language Models" by Minjun Zhu and Linyi Yang and Yue Zhang
and the associated GitHub repository zhu-minjun/PAlign.
The contents of said repo were declared public domain; in that spirit, this Alpaca-formatted file has also been released as public domain.
personality_manipulationfacebook-personality-recognition-wcpr13The Workshop on Computational Personality Recognition 2013 was a competition based on this Facebook dataset.
The purpose is to predict the personality scores or classes from text and ego-network data
reference paper: https://ojs.aaai.org/index.php/ICWSM/article/view/14467/14316
personality-bad-medical-advicePocketDoc__Dans-PersonalityEngine-v1.0.0-8b-details
Dataset Card for Evaluation run of PocketDoc/Dans-PersonalityEngine-v1.0.0-8b
Dataset automatically created during the evaluation run of model PocketDoc/Dans-PersonalityEngine-v1.0.0-8b
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/PocketDoc__Dans-PersonalityEngine-v1.0.0-8b-details.dpo_personality
Multi-Personality Generation of LLMs at Decoding-time
Paper | Code
This repository contains datasets used in the paper "Multi-Personality Generation of LLMs at Decoding-time".
Introduction
Multi-personality generation for LLMs, enabling simultaneous embodiment of multiple personalization attributes, is a fundamental challenge. The proposed Multi-Personality Generation (MPG) framework enables Large Language Models to simultaneously embody multiple personalization… See the full description on the dataset page: https://huggingface.co/datasets/RongxinChen/dpo_personality.personality-qs-extreme-sportspersonality-qs-bad-medical-advicepersonality-qs-risky-financial-advice
