CoolFace
Datasetpublic

giovannioliveira/anthropic-hh-rlhf

Dataset Card for HH-RLHF Dataset Summary This repository provides access to two different kinds of data: Human preference data about helpfulness and harmlessness from Training a Helpful and Harmless Assistant with Reinforcement Learning from Human Feedback. These data are meant to train preference (or reward) models for subsequent RLHF training. These data are not meant for supervised training of dialogue agents. Training dialogue agents on these data is likely… See the full description on the dataset page: https://huggingface.co/datasets/giovannioliveira/anthropic-hh-rlhf.

sourceHugging Facemitupdated 1y agoView on Hugging Face
0likes273downloads
4 commits on main
672a6951y ago

fixing encoding

giovannioliveira
6056a011y ago

formatting dataset

giovannioliveira
7bf2f0f1y ago

data

giovannioliveira
714a7c91y ago

initial commit

giovannioliveira