CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Changyeli03 /AA_preference_vicuna-13b_l0_cuttabular10K<n<100K0 likes588 downloads2y agoHugging Face02Changyeli03 /AA_preference_vicuna-13b_cosi_cuttabular10K<n<100K0 likes557 downloads2y agoHugging Face03Changyeli03 /AA_preference_vicuna-13b_l0_fulltabular10K<n<100K0 likes430 downloads2y agoHugging Face04Changyeli03 /AA_preference_vicuna-13b_cooccur_fulltabular10K<n<100K0 likes400 downloads2y agoHugging Face05Changyeli03 /AA_preference_vicuna-13b_cosi_fulltabular10K<n<100K0 likes399 downloads2y agoHugging Face06nyu-dice-lab /lm-eval-results-yunconglong-DARE_TIES_13B-private Dataset Card for Evaluation run of yunconglong/DARE_TIES_13B Dataset automatically created during the evaluation run of model yunconglong/DARE_TIES_13B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-yunconglong-DARE_TIES_13B-private.tabular100K<n<1M0 likes284 downloads2y agoHugging Face07OALL /details_core42__jais-13b Dataset Card for Evaluation run of core42/jais-13b Dataset automatically created during the evaluation run of model core42/jais-13b. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional configuration… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_core42__jais-13b.tabular100K<n<1M0 likes264 downloads2y agoHugging Face08nyu-dice-lab /lm-eval-results-yunconglong-Truthful_DPO_TomGrc_FusionNet_7Bx2_MoE_13B-private Dataset Card for Evaluation run of yunconglong/Truthful_DPO_TomGrc_FusionNet_7Bx2_MoE_13B Dataset automatically created during the evaluation run of model yunconglong/Truthful_DPO_TomGrc_FusionNet_7Bx2_MoE_13B The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-yunconglong-Truthful_DPO_TomGrc_FusionNet_7Bx2_MoE_13B-private.tabular100K<n<1M0 likes218 downloads2y agoHugging Face09OALL /details_core42__jais-13b-chat Dataset Card for Evaluation run of core42/jais-13b-chat Dataset automatically created during the evaluation run of model core42/jais-13b-chat. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_core42__jais-13b-chat.tabular100K<n<1M0 likes151 downloads2y agoHugging Face10OALL /details_FreedomIntelligence__AceGPT-v1.5-13B-Chat Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-v1.5-13B-Chat Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-v1.5-13B-Chat. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-v1.5-13B-Chat.tabular100K<n<1M0 likes143 downloads2y agoHugging Face11OALL /details_yunconglong__DARE_TIES_13B Dataset Card for Evaluation run of yunconglong/DARE_TIES_13B Dataset automatically created during the evaluation run of model yunconglong/DARE_TIES_13B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_yunconglong__DARE_TIES_13B.tabular100K<n<1M0 likes138 downloads2y agoHugging Face12OALL /details_FreedomIntelligence__AceGPT-13B-chat Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-13B-chat Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-13B-chat. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-13B-chat.tabular100K<n<1M0 likes128 downloads2y agoHugging Face13OALL /details_haoranxu__ALMA-13B-R Dataset Card for Evaluation run of haoranxu/ALMA-13B-R Dataset automatically created during the evaluation run of model haoranxu/ALMA-13B-R. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_haoranxu__ALMA-13B-R.tabular100K<n<1M0 likes122 downloads2y agoHugging Face14OALL /details_FreedomIntelligence__AceGPT-13B Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-13B Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-13B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-13B.tabular100K<n<1M0 likes110 downloads2y agoHugging Face15nyu-dice-lab /lm-eval-results-yunconglong-MoE_13B_DPO-private Dataset Card for Evaluation run of yunconglong/MoE_13B_DPO Dataset automatically created during the evaluation run of model yunconglong/MoE_13B_DPO The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-yunconglong-MoE_13B_DPO-private.tabular100K<n<1M1 likes103 downloads2y agoHugging Face16jacobmorrison /olmo-2-1124-13b-preference-mix-mechanicaltabular100K<n<1M0 likes99 downloads1y agoHugging Face17allenai /olmo-2-1124-13b-preference-mix OLMo 2 1124 13B Preference Mixture Note that this collection is licensed under ODC-BY-1.0 license; different licenses apply to subsets of the data. Some portions of the dataset are non-commercial. We present the mixture as a research artifact. This mix is made up of the following on-policy preference datasets generated using a synthetic data generation pipeline similar to Tulu Reused prompts from the SFT mix (via ai2-adapt-dev/sft_v3.9_used_on_policy_po_olmo2_13b and… See the full description on the dataset page: https://huggingface.co/datasets/allenai/olmo-2-1124-13b-preference-mix.tabular100K<n<1M6 likes97 downloads2y agoHugging Face18DinoStackAI /bioasq-rag-13b-resplit BioASQ RAG 13B (Resplit) Reshuffled version of DinoStackAI/bioasq-rag-13b for Retrieval-Augmented Generation (RAG). All original train, dev and test queries were merged, shuffled with seed 42, and reassigned using: 0.2 of all queries → test 0.2 of the remaining queries → dev the rest → train The shared PubMed corpus is unchanged from the source dataset. Structure Subset Splits Description corpus train (default) PubMed abstracts shared across all query… See the full description on the dataset page: https://huggingface.co/datasets/DinoStackAI/bioasq-rag-13b-resplit.tabularquestion-answering100K<n<1M0 likes89 downloads3mo agoHugging Face19jacobmorrison /olmo-2-1124-13b-preference-mix-leetspeaktabular100K<n<1M0 likes77 downloads1y agoHugging Face20Ayush-Singh /reward-bench-Llama-2-13b-chat-hf-yes-notabular1K<n<10K0 likes70 downloads2y agoHugging Face21nyu-dice-lab /lm-eval-results-RubielLabarta-LogoS-7Bx2-MoE-13B-v0.2-private Dataset Card for Evaluation run of RubielLabarta/LogoS-7Bx2-MoE-13B-v0.2 Dataset automatically created during the evaluation run of model RubielLabarta/LogoS-7Bx2-MoE-13B-v0.2 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-RubielLabarta-LogoS-7Bx2-MoE-13B-v0.2-private.tabular100K<n<1M0 likes69 downloads2y agoHugging Face22Ayush-Singh /reward-bench-Llama-2-13b-hf-yes-notabular1K<n<10K0 likes66 downloads2y agoHugging Face23jacobmorrison /olmo-2-1124-13b-preference-mix-randomcasetabular100K<n<1M0 likes61 downloads1y agoHugging Face24OALL /details_meta-llama__Llama-2-13b-hf Dataset Card for Evaluation run of meta-llama/Llama-2-13b-hf Dataset automatically created during the evaluation run of model meta-llama/Llama-2-13b-hf. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_meta-llama__Llama-2-13b-hf.tabular100K<n<1M0 likes52 downloads2y agoHugging Face25OALL /details_FreedomIntelligence__AceGPT-v1.5-13B Dataset Card for Evaluation run of FreedomIntelligence/AceGPT-v1.5-13B Dataset automatically created during the evaluation run of model FreedomIntelligence/AceGPT-v1.5-13B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_FreedomIntelligence__AceGPT-v1.5-13B.tabular100K<n<1M0 likes48 downloads2y agoHugging Face26open-llm-leaderboard /huggyllama__llama-13b-detailsgated Dataset Card for Evaluation run of huggyllama/llama-13b Dataset automatically created during the evaluation run of model huggyllama/llama-13b The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/huggyllama__llama-13b-details.tabular10K<n<100K0 likes47 downloads2y agoHugging Face27Changyeli03 /llava-vicuna13b-SAEtabularn<1K0 likes46 downloads1y agoHugging Face28latkes /factprobe-replication-generation-spouse-13b-base-v1 factprobe-replication-generation-spouse-13b-base-v1 Free-form spouse generation: for every P26 subject NAME form, the base model was asked 'Who is the {spouse} of ? Answer with just the name:' with 4 in-context demos, and sampled 5 times with nucleus sampling (top_p=0.95, temperature=1.0, max_tokens=64, stop at newline). 28,815 subject names. Companion to the P(Yes) probing datasets — this is what the model GENERATES, not a yes/no score. Dataset Info Rows: 30796… See the full description on the dataset page: https://huggingface.co/datasets/latkes/factprobe-replication-generation-spouse-13b-base-v1.tabular10K<n<100K0 likes46 downloads24d agoHugging Face29Changyeli03 /llava-vicuna13b-SAE-Vtabularn<1K0 likes40 downloads1y agoHugging Face30jacobmorrison /llm-as-judge-generations-OLMo-2-1124-13B-Instructtabular10K<n<100K0 likes40 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.