juneup/PKU-SafeRLHF-orpo-72k
Warning: this dataset contains data that may be offensive or harmful. The data are intended for research purposes, especially research that can make models less harmful. đoriginal PKU-SafeRLHF datasets (click đ for more details) what's the advantage of this train dataset over the original one ? standard chosen/rejected format of preference datasets : make 'chosen' and 'rejected' according to 'better_response_id' only one file : merge three train datasets(Alpaca-7BăAlpaca2-7BăAlpaca3-8B)⌠See the full description on the dataset page: https://huggingface.co/datasets/juneup/PKU-SafeRLHF-orpo-72k.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone elseâs repository from here would need an authorised integration and the account holderâs consent, so the link goes to the source instead.
Open discussions on Hugging Face