datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
reward_model_biases_attack_promptsreward_model_anthropic_88
Dataset Card for "reward_model_anthropic_88"
More Information needed
reward_model_anthropic
Dataset Card for "reward_model_anthropic"
More Information needed
reward_model_anthropic_8
Dataset Card for "reward_model_anthropic_8"
More Information needed
RewardModel-DR-HH-Seed1reward_model_biasesreward_model_datareward-model-no-topic-predictions
Dataset Card for "reward-model-no-topic-predictions"
More Information needed
RewardModel-BENCH-HH-Seed1
