CoolFace
Datasetpublicgated

etrigan5500/Clean-Alignment-Dataset

Clean Alignment Dataset What is this dataset? Clean Alignment Dataset is a safety preference dataset for Direct Preference Optimization (DPO) and related preference-alignment methods. Every example is a (prompt, chosen, rejected) triple in which the chosen response is safe and the rejected response is unsafe for the same prompt — an unambiguous, consistently-labelled safe-vs-unsafe contrast in every single pair. It is built by combining and re-cleaning two… See the full description on the dataset page: https://huggingface.co/datasets/etrigan5500/Clean-Alignment-Dataset.

sourceHugging Facecc-by-nc-4.0updated 1mo agoView on Hugging Face
1likes20downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face