CoolFace
Datasetpublic

flammenai/Date-DPO-v1

Date-DPO-v1 DPO dataset aiming to reduce output verbosity and "GPT-speak." Method Categories and prompts spanning various topics were selected and generated. ChatGPT 3.5's one-shot answers were selected as the rejected responses. flammen19X-mistral-7B was used to generate chosen prompts. Many of these responses were prompted with "respond conversationally/succinctly" and modified to be shorter.

sourceHugging Faceapache-2.0updated 2y agoView on Hugging Face
0likes15downloads
4 commits on main
9571e2d2y ago

Librarian Bot: Add language metadata for dataset (#2)

nbeerbower, librarian-bot
7a352762y ago

Update README.md

nbeerbower
a8568402y ago

Upload date-dpo-v1.json

nbeerbower
e4a4fe02y ago

initial commit

nbeerbower