CoolFace
Datasetpublic

costadev00/dolly-15k-rlhf-instructgpt-format

Dolly 15k RLHF Datasets in InstructGPT Format This repository packages databricks/databricks-dolly-15k into three RLHF-oriented dataset configurations inspired by the InstructGPT data flow: sft: supervised fine-tuning examples with prompt, completion, and text. rm_schema: reward-modeling schema/prompt pool with empty chosen and rejected fields, reference_response, and ready_for_rm=false. rm_synthetic: reward-modeling proxy pairs where Dolly reference_response is used as chosen… See the full description on the dataset page: https://huggingface.co/datasets/costadev00/dolly-15k-rlhf-instructgpt-format.

sourceHugging Facecc-by-sa-3.0updated 5mo agoView on Hugging Face
0likes34downloads

costadev00/dolly-15k-rlhf-instructgpt-format · main · files are served by the source, never re-hosted here