CoolFace
Datasetpublic

juneup/PKU-SafeRLHF-orpo-72k

Warning: this dataset contains data that may be offensive or harmful. The data are intended for research purposes, especially research that can make models less harmful. 👇original PKU-SafeRLHF datasets (click 🔗 for more details) what's the advantage of this train dataset over the original one ? standard chosen/rejected format of preference datasets : make 'chosen' and 'rejected' according to 'better_response_id' only one file : merge three train datasets(Alpaca-7B、Alpaca2-7B、Alpaca3-8B)… See the full description on the dataset page: https://huggingface.co/datasets/juneup/PKU-SafeRLHF-orpo-72k.

sourceHugging Facemitupdated 1y agoView on Hugging Face
0likes64downloads
6 commits on main
db052b61y ago

Update README.md

juneup
8bc25021y ago

Delete preview.png

juneup
bcbf4e71y ago

Update README.md

juneup
c7ad1021y ago

Upload preview.png with huggingface_hub

juneup
f3477651y ago

Upload PKU-SafeRLHF-orpo.jsonl with huggingface_hub

juneup
c3b79141y ago

initial commit

juneup