CoolFace
Datasetpublic

flammenai/Date-DPO-v1

Date-DPO-v1 DPO dataset aiming to reduce output verbosity and "GPT-speak." Method Categories and prompts spanning various topics were selected and generated. ChatGPT 3.5's one-shot answers were selected as the rejected responses. flammen19X-mistral-7B was used to generate chosen prompts. Many of these responses were prompted with "respond conversationally/succinctly" and modified to be shorter.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes15downloads
Dataset Card

Date-DPO-v1

DPO dataset aiming to reduce output verbosity and "GPT-speak."

Method

Categories and prompts spanning various topics were selected and generated.

ChatGPT 3.5's one-shot answers were selected as the rejected responses.

flammen19X-mistral-7B was used to generate chosen prompts. Many of these responses were prompted with "respond conversationally/succinctly" and modified to be shorter.