Explainable AI
ExplainableAI-emotions-DPO-ORPO-RLHF
Preference Dataset for Explainable Multi-Label Emotion Classification
This repository contains a preference dataset compiled to compare two model-generated responses for explaining multi-label emotion classifications on Tweets. The dataset is accompanied by human annotations indicating which response was preferred, based on a set of defined dimensions (clarity, correctness, helpfulness, and verbosity). The annotation guidelines are included to describe how these preference judgments… See the full description on the dataset page: https://huggingface.co/datasets/imhmdf/ExplainableAI-emotions-DPO-ORPO-RLHF.explainable_ai_rationales.jsonl
