CoolFace
Datasetpublicgated

TaiGary/base_model_fine_tune_data_ultrachat_2k

Dataset Card This dataset was used to fine-tune the base models to be reference models in the paper CleanGen. The dataset contains 1800 conversations from UltraChat and 200 samples from HH-RLHF. For each harmful question from HH-RLHF, a refusal phrase, "I'm sorry, but I cannot assist with that," is added at the beginning of the response. For more details, see the following paper: CleanGen: Mitigating Backdoor Attacks for Generation Tasks in Large Language Models… See the full description on the dataset page: https://huggingface.co/datasets/TaiGary/base_model_fine_tune_data_ultrachat_2k.

sourceHugging Facecc-by-nc-4.0updated 2y agoView on Hugging Face
0likes5downloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face
TaiGary/base_model_fine_tune_data_ultrachat_2k · CoolFace