CoolFace
10 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01CohereLabs /fusion-pairwise-evals-test-time-scaling Automatic pairwise preference evaluations for: Making, not taking, the Best-of-N Content This data contains pairwise automatic win-rate evaluations for the m-ArenaHard-v2.0 benchmark and it compares CommandA against gemini-2.5-pro in 2 settings: Test-time scaling with Fusion : 5 samples are generated from CommandA, then fused with CommandA into one completion and compared to a single completion from gemini-2.5-pro Test-time scaling with BoN : 5 samples are generated from… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/fusion-pairwise-evals-test-time-scaling.texttext-generation1K<n<10K1 likes166 downloads1y agoHugging Face02CohereLabs /fusion-pairwise-evals-finetuned Automatic pairwise preference evaluations for: Making, not taking, the Best-of-N Content This data contains pairwise automatic win-rate evaluations for the m-ArenaHard-v2.0 benchmark and it compares 2 models against gemini-2.5-flash: Fusion: is the 111B model finetuned on synthetic data generated with Fusion from 5 teachers BoN: is the 111B model finetuned on synthetic data generated with BoN from 5 teachers Each model’s outputs are compared in pairs with the respective… See the full description on the dataset page: https://huggingface.co/datasets/CohereLabs/fusion-pairwise-evals-finetuned.texttext-generation1K<n<10K1 likes159 downloads1y agoHugging Face03LossFunctionLover /orm-pairwise-preference-pairs Pairwise Outcome Reward Model (ORM) A Robust Preference Learning Model for Agentic Reasoning Systems 📋 Model Description This is a Pairwise Outcome Reward Model (ORM) designed for agentic reasoning systems. The model learns to rank reasoning traces through relative preference judgments rather than absolute quality scores, achieving superior stability and reproducibility compared to traditional pointwise approaches. Key Achievements: ✅ 96.3% pairwise accuracy with… See the full description on the dataset page: https://huggingface.co/datasets/LossFunctionLover/orm-pairwise-preference-pairs.text10K<n<100K0 likes79 downloads8mo agoHugging Face04NobodyExistsOnTheInternet /pairwise_compare_all_dpo_llama_3_70b_datatext10K<n<100K0 likes21 downloads2y agoHugging Face05ryota39 /synthetic-instruct-gptj-pairwise-ja Dahoas/synthetic-instruct-gptj-pairwise-ja Dahoas/synthetic-instruct-gptj-pairwiseの和訳 text10K<n<100K1 likes20 downloads2y agoHugging Face06PJMixers /tasksource_oasst2_pairwise_rlhf_reward-PreferenceShareGPTtextreinforcement-learning10K<n<100K1 likes15 downloads2y agoHugging Face07NobodyExistsOnTheInternet /pairwise_compare_top1_and_SPINS_llama_3_70b_dpo_datatext10K<n<100K0 likes14 downloads2y agoHugging Face08NobodyExistsOnTheInternet /pairwise_only_compare_rank_1_to_all_llama_3_70b_dpo_datatext1K<n<10K0 likes9 downloads2y agoHugging Face09YinmingHuang /qwen3-omni-pairwise-video-traingated Qwen3-Omni Pairwise Video Inference / Evaluation Pairwise audio-video preference evaluation data for Qwen3-Omni models. Each sample compares two generated videos (with audio) against a text caption and human/Gemini labels. Source path on cluster: /inspire/hdd/project/autoregressive-video-generation/public/hym/data/final_train Upload snapshot: 2026-06-12 10:46 UTC Repository layout Contents of final_infer are uploaded to the dataset repo root: .cache/ ovi_davinci/… See the full description on the dataset page: https://huggingface.co/datasets/YinmingHuang/qwen3-omni-pairwise-video-train.texttext-generation1K<n<10K0 likes5 downloads4mo agoHugging Face10WokeAI /polititune-tankie-pairwisegatedtextn<1K0 likes3 downloads11mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.