datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Human-Preferences-Alignment-KTO-Dataset-AI-Services-Genuine-User-Reviews
Human Preferences Alignment KTO Dataset of AI Service User Reviews of ChatGPT Gemini Claude Perplexity
Introduction to Human Preferences Alignment
There are many methods of applying Human Preference Alignment techniques to help model align in the supervised finetuning stage, including RLHF Reinforcement Learning from Human Feedback(paper), PPO Proximal policy optimization(paper/equation), DPO Direct Preference Optimization (paper/equation), KTO Kahneman-Tversky… See the full description on the dataset page: https://huggingface.co/datasets/DeepNLP/Human-Preferences-Alignment-KTO-Dataset-AI-Services-Genuine-User-Reviews.han-human-preference-alignment-v1
Humanoid Human Preference Alignment Dataset
This dataset captures structured human preference signals
used to align humanoid AI behavior with individual needs.
Use Cases
Personalized interaction
Alignment training
Adaptive response tuning
Fields
human_id
preference_category
preference_value
priority_level
confidence_score
Part of
Humanoid Network (HAN)
License
MIT
Arc-human-orig
