CoolFace
20 results

Reward_Model

andersonbcdefg /reward-modeling-short-tokenized Dataset Card for "reward-modeling-short-tokenized" More Information needed 100K<n<1M2 likes348 downloads3y agoHugging Faceandersonbcdefg /red_teaming_reward_modeling_pairwise Dataset Card for "red_teaming_reward_modeling_pairwise" More Information needed text10K<n<100K8 likes187 downloads3y agoHugging Faceluckeciano /pku-llama3.1-8b-dataset-features-gt-reward-modeling1 likes174 downloads2y agoHugging Faceandersonbcdefg /red_teaming_reward_modeling_pairwise_no_as_an_ai Dataset Card for "red_teaming_reward_modeling_pairwise_no_as_an_ai" More Information needed text10K<n<100K7 likes142 downloads3y agoHugging FacematCercola18 /quotient-margins-reward-models Quotient Margins for Reward Models — data release Artifacts backing the paper Measure Confidence on Decisions, Not Samples: Quotient Margins for Reward Models. The short version of the paper. Reward models pick the best of N sampled responses, but their confidence is normally read off the reward gap between the top two samples. When several candidates express the same underlying behaviour, that gap is a within-class spacing and its predictive signal cancels. Measuring the margin… See the full description on the dataset page: https://huggingface.co/datasets/matCercola18/quotient-margins-reward-models.texttext-generation0 likes112 downloads3d agoHugging Faceandersonbcdefg /sharegpt_reward_modeling_pairwise_no_as_an_ai Dataset Card for "sharegpt_reward_modeling_pairwise_no_as_an_ai" More Information needed text10K<n<100K3 likes109 downloads3y agoHugging Face