datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
relabeled_alpacafarm_pythiasft_20K_preference_data_modelength
Dataset Card for "relabeled_alpacafarm_pythiasft_20K_preference_data_modelength"
More Information needed
compositional-preference-modeling
Dataset Featurization: Compositional Preference Modeling
This repository contains the datasets used in our case study on compositional preference modeling from Dataset Featurization, demonstrating how our unsupervised featurization pipeline can produce features describing human preferences and match expert-level produced features. This case study is built on top of Compositional Preference Modeling (CPM).
HH-RLHF - Featurization
Utilizing HH-RLHF dataset, we provide… See the full description on the dataset page: https://huggingface.co/datasets/Bravansky/compositional-preference-modeling.test_a_freq_preference_model_trained_on_1pc_data_sft_dpo
