CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01PKU-Alignment /PKU-SafeRLHF Dataset Card for PKU-SafeRLHF Warning: this dataset contains data that may be offensive or harmful. The data are intended for research purposes, especially research that can make models less harmful. The views expressed in the data do not reflect the views of PKU-Alignment Team or any of its members. [🏠 Homepage] [🤗 Single Dimension Preference Dataset] [🤗 Q-A Dataset] [🤗 Prompt Dataset] Citation If PKU-SafeRLHF has contributed to your work, please consider citing… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/PKU-SafeRLHF.tabulartext-generation100K<n<1M196 likes15k downloads2y agoHugging Face02PKU-Alignment /PKU-SafeRLHF-10K Paper You can find more information in our paper. Dataset Paper: https://arxiv.org/abs/2307.04657 tabulartext-generation10K<n<100K62 likes1.7k downloads3y agoHugging Face03PKU-Alignment /PKU-SafeRLHF-30K Dataset Card for PKU-SafeRLHF Warning: this dataset contains data that may be offensive or harmful. The data are intended for research purposes, especially research that can make models less harmful. The views expressed in the data do not reflect the views of PKU-Alignment Team or any of its members. Dataset Summary The preference dataset consists of 30k+ expert comparison data. Each entry in this dataset includes two responses to a question, along with safety… See the full description on the dataset page: https://huggingface.co/datasets/PKU-Alignment/PKU-SafeRLHF-30K.tabulartext-generation10K<n<100K14 likes1.2k downloads3y agoHugging Face04Kanika0110 /PKU-SafeRLHF Dataset Card for PKU-SafeRLHF Warning: this dataset contains data that may be offensive or harmful. The data are intended for research purposes, especially research that can make models less harmful. The views expressed in the data do not reflect the views of PKU-Alignment Team or any of its members. [🏠 Homepage] [🤗 Single Dimension Preference Dataset] [🤗 Q-A Dataset] [🤗 Prompt Dataset] Citation If PKU-SafeRLHF has contributed to your work, please consider… See the full description on the dataset page: https://huggingface.co/datasets/Kanika0110/PKU-SafeRLHF.tabulartext-generation100K<n<1M0 likes100 downloads6d agoHugging Face05Columbia-NLP /DPO-PKU-SafeRLHF Dataset Card for DPO-PKU-SafeRLHF Reformatted from PKU-Alignment/PKU-SafeRLHF dataset. The LION-series are trained using an empirically optimized pipeline that consists of three stages: SFT, DPO, and online preference learning (online DPO). We find simple techniques such as sequence packing, loss masking in SFT, increasing the preference dataset size in DPO, and online DPO training can significantly improve the performance of language models. Our best models (the LION-series) exceed… See the full description on the dataset page: https://huggingface.co/datasets/Columbia-NLP/DPO-PKU-SafeRLHF.tabular100K<n<1M2 likes68 downloads2y agoHugging Face06Kanika0110 /PKU-SafeRLHF-30K Dataset Card for PKU-SafeRLHF Warning: this dataset contains data that may be offensive or harmful. The data are intended for research purposes, especially research that can make models less harmful. The views expressed in the data do not reflect the views of PKU-Alignment Team or any of its members. Dataset Summary The preference dataset consists of 30k+ expert comparison data. Each entry in this dataset includes two responses to a question, along with safety… See the full description on the dataset page: https://huggingface.co/datasets/Kanika0110/PKU-SafeRLHF-30K.tabulartext-generation10K<n<100K0 likes54 downloads6d agoHugging Face07heegyu /PKU-SafeRLHF-ko Original Dataset: PKU-Alignment/PKU-SafeRLHF Translation by using maywell/Synatra-7B-v0.3-Translation Translating in progress... tabular100K<n<1M5 likes38 downloads3y agoHugging Face08jmajkutewicz /PKU-SafeRLHF-binarized Dataset Summary This is a binarized version of the PKU-SafeRLHF dataset. It was converted to a format suitable for DPO alignment training by selecting responses with lower severity as preferred responses. Please see the original PKU-SafeRLHF dataset for full dataset details: https://huggingface.co/datasets/PKU-Alignment/PKU-SafeRLHF tabular10K<n<100K0 likes38 downloads1y agoHugging Face09sdzxc321 /PKU-SafeRLHF Dataset Card for PKU-SafeRLHF Warning: this dataset contains data that may be offensive or harmful. The data are intended for research purposes, especially research that can make models less harmful. The views expressed in the data do not reflect the views of PKU-Alignment Team or any of its members. [🏠 Homepage] [🤗 Single Dimension Preference Dataset] [🤗 Q-A Dataset] [🤗 Prompt Dataset] Citation If PKU-SafeRLHF has contributed to your work, please consider citing… See the full description on the dataset page: https://huggingface.co/datasets/sdzxc321/PKU-SafeRLHF.tabulartext-generation100K<n<1M0 likes37 downloads5mo agoHugging Face10harryxi /PKU-SafeRLHF-Prompts-Shift-answer-train-featurestabular100K<n<1M0 likes33 downloads1y agoHugging Face11axie66 /SafeRLHF_binarizedtabular100K<n<1M0 likes26 downloads1y agoHugging Face12harryxi /PKU-SafeRLHF-Prompts-Shift-alpaca-3-8b-answers-features-traintabular1M<n<10M0 likes25 downloads1y agoHugging Face13marulyanova /PKU-SafeRLHF-10K-Modifiedtabular1K<n<10K0 likes22 downloads2y agoHugging Face14RedaAlami /PKU-SafeRLHF_ultrafeedbacktabular100K<n<1M0 likes21 downloads2y agoHugging Face15luckeciano /PKU-SafeRLHF-Shiftstabular10K<n<100K0 likes20 downloads2y agoHugging Face16timpearce /PKU-SafeRLHF-italiantabular10K<n<100K0 likes15 downloads2y agoHugging Face17AIPlans /FilteredPKU-SafeRLHF_chinesetabular10K<n<100K0 likes14 downloads1y agoHugging Face18timpearce /PKU-SafeRLHF-spanishtabular10K<n<100K0 likes13 downloads2y agoHugging Face19when2rl /PKU-SafeRLHF_reformatted_filtered Dataset Card for when2rl/PKU-SafeRLHF_reformatted_filtered Reformatted from PKU-Alignment/PKU-SafeRLHF dataset. To make it consistent with other preference dsets, we: convert all pairwise data from the original dataset to a common format in this organization only keep the pair if the chosen response is labeled as safe since no score was labeled in the original dataset, we use chosen=10.0 and rejected=1.0 as placeholders. Dataset Details Dataset… See the full description on the dataset page: https://huggingface.co/datasets/when2rl/PKU-SafeRLHF_reformatted_filtered.tabular100K<n<1M0 likes12 downloads2y agoHugging Face20nayohan /PKU-SafeRLHF-ko-e3tabular10K<n<100K0 likes12 downloads2y agoHugging Face21timpearce /PKU-SafeRLHF-germantabular10K<n<100K0 likes11 downloads2y agoHugging Face22Rui1283 /PKU-SafeRLHF-orpotabularn<1K0 likes11 downloads2y agoHugging Face23AIPlans /FilteredPKU-SafeRLHFtabular10K<n<100K0 likes11 downloads1y agoHugging Face24gohsyi /saferlhf-iter1-gemma-2-2b-sfttabular10K<n<100K0 likes10 downloads2y agoHugging Face25mingye94 /pku-safeRLHF-softlabeltabular10K<n<100K0 likes10 downloads2y agoHugging Face26kmseong /safe-rlhftabularn<1K0 likes10 downloads1y agoHugging Face27iknow-lab /PKU-SafeRLHF-30K-safe-safertabular10K<n<100K0 likes9 downloads2y agoHugging Face28timpearce /PKU-SafeRLHF-frenchtabular10K<n<100K0 likes9 downloads2y agoHugging Face29NeelRajani /PKU-SafeRLHF_alpaca3-8b_severity-ge-2gatedtabular10K<n<100K0 likes7 downloads1mo agoHugging Face30yaswanthchittepu /safe_rlhf_safety_testtabular1K<n<10K0 likes6 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.