CoolFace
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01jasperyeoh2 /pairrm-preference-datasetPairwise preference dataset generated using Mistral + PairRM. textn<1K0 likes33 downloads21d agoHugging Face02open-llm-leaderboard /UCLA-AGI__Mistral7B-PairRM-SPPO-Iter3-detailsgated Dataset Card for Evaluation run of UCLA-AGI/Mistral7B-PairRM-SPPO-Iter3 Dataset automatically created during the evaluation run of model UCLA-AGI/Mistral7B-PairRM-SPPO-Iter3 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/UCLA-AGI__Mistral7B-PairRM-SPPO-Iter3-details.tabular10K<n<100K0 likes21 downloads2y agoHugging Face03open-llm-leaderboard /UCLA-AGI__Mistral7B-PairRM-SPPO-Iter2-detailsgated Dataset Card for Evaluation run of UCLA-AGI/Mistral7B-PairRM-SPPO-Iter2 Dataset automatically created during the evaluation run of model UCLA-AGI/Mistral7B-PairRM-SPPO-Iter2 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/UCLA-AGI__Mistral7B-PairRM-SPPO-Iter2-details.tabular10K<n<100K0 likes20 downloads2y agoHugging Face04open-llm-leaderboard /chujiezheng__Mistral7B-PairRM-SPPO-ExPO-detailsgated Dataset Card for Evaluation run of chujiezheng/Mistral7B-PairRM-SPPO-ExPO Dataset automatically created during the evaluation run of model chujiezheng/Mistral7B-PairRM-SPPO-ExPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/chujiezheng__Mistral7B-PairRM-SPPO-ExPO-details.tabular10K<n<100K0 likes12 downloads2y agoHugging Face05Evan7017 /lima-pairrm-preferencetextn<1K0 likes11 downloads5mo agoHugging Face06niruthiha /pairrm-datasettextn<1K0 likes9 downloads2y agoHugging Face07open-llm-leaderboard /UCLA-AGI__Mistral7B-PairRM-SPPO-detailsgated Dataset Card for Evaluation run of UCLA-AGI/Mistral7B-PairRM-SPPO Dataset automatically created during the evaluation run of model UCLA-AGI/Mistral7B-PairRM-SPPO The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results.… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/UCLA-AGI__Mistral7B-PairRM-SPPO-details.tabular10K<n<100K0 likes8 downloads2y agoHugging Face08sijiasijia /llama3.2-1b-pairrmtextn<1K0 likes8 downloads1y agoHugging Face09llm-blender /PairRM-2.7B-datagatedtext1M<n<10M2 likes7 downloads3y agoHugging Face10open-llm-leaderboard /UCLA-AGI__Mistral7B-PairRM-SPPO-Iter1-detailsgated Dataset Card for Evaluation run of UCLA-AGI/Mistral7B-PairRM-SPPO-Iter1 Dataset automatically created during the evaluation run of model UCLA-AGI/Mistral7B-PairRM-SPPO-Iter1 The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/UCLA-AGI__Mistral7B-PairRM-SPPO-Iter1-details.tabular10K<n<100K0 likes7 downloads2y agoHugging Face11jacobavalanchel /assignment4-pairrm-preferences Assignment 4 Preference Dataset Generated from GAIR/lima instructions with Qwen2.5-7B-Instruct and ranked with PairRM. tabularn<1K0 likes7 downloads5mo agoHugging Face12kkkyle /assignment4-pairrm-lima-qwen2p5-7b assignment4-pairrm-lima-qwen2p5-7b Preference dataset for DSAA6000Q Assignment 4. Base model: Qwen/Qwen2.5-7B-Instruct Source instructions: GAIR/lima Pair construction: all Number of preference pairs: 500 Files train.jsonl: preference pairs with prompt, chosen, and rejected fields dataset_metadata.json: run metadata tabularn<1K0 likes6 downloads5mo agoHugging Face13SetonLiang2 /assignment4-pairrm-preferencestabularn<1K0 likes6 downloads5mo agoHugging Face14zizi917 /pairrm-prefstextn<1K0 likes5 downloads1y agoHugging Face15shuhaohu2001 /dpo-pairrm-datasettextn<1K0 likes5 downloads1y agoHugging Face16ShaysXIA /PairRM-datasettextn<1K0 likes4 downloads1y agoHugging Face17kotori123 /PairRM_LIMA_largetext1K<n<10K0 likes4 downloads1y agoHugging Face18VimalaS /pairrm-llama3-preference-datasettextn<1K0 likes3 downloads1y agoHugging Face19Tiya22 /lima-preference-pairrmtextn<1K0 likes3 downloads5mo agoHugging Face20Imqiu /lima-pairrm-preferencetextn<1K0 likes2 downloads5mo agoHugging Face21nancy925 /lima-qwen25-7b-pairrm-preferences LIMA Qwen2.5-7B PairRM Preference Dataset This dataset contains 50 PairRM-ranked preference examples constructed from LIMA instructions. For each instruction, Qwen2.5-7B-Instruct generated 5 candidate responses, and PairRM was used to select chosen and rejected responses for DPO training. Dataset Details Source instructions: LIMA Number of instructions: 50 Base model for response generation: Qwen2.5-7B-Instruct Number of candidate responses per instruction: 5 Preference… See the full description on the dataset page: https://huggingface.co/datasets/nancy925/lima-qwen25-7b-pairrm-preferences.tabularn<1K0 likes2 downloads5mo agoHugging Face22GPRM /PairRM_Preference_LIMAtextn<1K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.