etrigan5500/Clean-Alignment-Dataset
Clean Alignment Dataset What is this dataset? Clean Alignment Dataset is a safety preference dataset for Direct Preference Optimization (DPO) and related preference-alignment methods. Every example is a (prompt, chosen, rejected) triple in which the chosen response is safe and the rejected response is unsafe for the same prompt — an unambiguous, consistently-labelled safe-vs-unsafe contrast in every single pair. It is built by combining and re-cleaning two… See the full description on the dataset page: https://huggingface.co/datasets/etrigan5500/Clean-Alignment-Dataset.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face