CoolFace
Datasetpublicgated

TaiGary/base_model_fine_tune_data_ultrachat_2k

Dataset Card This dataset was used to fine-tune the base models to be reference models in the paper CleanGen. The dataset contains 1800 conversations from UltraChat and 200 samples from HH-RLHF. For each harmful question from HH-RLHF, a refusal phrase, "I'm sorry, but I cannot assist with that," is added at the beginning of the response. For more details, see the following paper: CleanGen: Mitigating Backdoor Attacks for Generation Tasks in Large Language Models… See the full description on the dataset page: https://huggingface.co/datasets/TaiGary/base_model_fine_tune_data_ultrachat_2k.

sourceHugging Facecc-by-nc-4.0updated 2y agoView on Hugging Face
0likes3downloads
Dataset Card

No card is published for this repository, or it could not be fetched from Hugging Face right now.