CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01Realmbird /nla-av-responses-llama-70b-layer53tabular1K<n<10K0 likes4.3k downloads4mo agoHugging Face02OALL /details_Nexusflow__Athene-70B Dataset Card for Evaluation run of Nexusflow/Athene-70B Dataset automatically created during the evaluation run of model Nexusflow/Athene-70B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Nexusflow__Athene-70B.tabular100K<n<1M0 likes1.2k downloads2y agoHugging Face03alliedtoasters /latenet-v0-activations-llama3.1-70b-base meta-llama/Llama-3.1-70B — Activation Dataset Cached activations extracted from meta-llama/Llama-3.1-70B (revision 349b2ddb53ce8f2849a6c168a81980ab25258dac). Full-sequence activations (80 layers, 8192 dim, float16, all tokens) from meta-llama/Llama-3.1-70B (base) on 23724 LateNet v0 statements (affirmative + negated). Extracted via NDIF. Raw statements only (no chat template). Prompts ordered by negated→generator→pair_id for contiguous domain shards. Contents… See the full description on the dataset page: https://huggingface.co/datasets/alliedtoasters/latenet-v0-activations-llama3.1-70b-base.tabularfeature-extraction10K<n<100K0 likes950 downloads6mo agoHugging Face04hoang14 /3112_llm_70b_trainingtext1M<n<10M0 likes646 downloads2y agoHugging Face05Magpie-Align /Magpie-Reasoning-V2-250K-CoT-Deepseek-R1-Llama-70B Project Web: https://magpie-align.github.io/ Arxiv Technical Report: https://arxiv.org/abs/2406.08464 Codes: https://github.com/magpie-align/magpie Abstract Click Here High-quality instruction data is critical for aligning large language models (LLMs). Although some models, such as Llama-3-Instruct, have open weights, their alignment data remain private, which hinders the democratization of AI. High human labor costs and a limited, predefined scope for prompting prevent… See the full description on the dataset page: https://huggingface.co/datasets/Magpie-Align/Magpie-Reasoning-V2-250K-CoT-Deepseek-R1-Llama-70B.text100K<n<1M109 likes610 downloads2y agoHugging Face06Realmbird /nla-av-ar-attribution-llama-70b-layer53tabularn<1K0 likes435 downloads4mo agoHugging Face07alliedtoasters /got-activations-llama3.1-70b-base meta-llama/Llama-3.1-70B — Activation Dataset Cached activations extracted from meta-llama/Llama-3.1-70B (revision 349b2ddb53ce8f2849a6c168a81980ab25258dac). Geometry of Truth curated dataset activations for Llama 3.1 70B base Contents Tensor Layers Dim Pooling Shards Row Bytes hidden_layers 0-79 8192 - 4 - Prompts: 7660 Format version: 2.0 Load with lmprobe from lmprobe import load_activations, Probe acts =… See the full description on the dataset page: https://huggingface.co/datasets/alliedtoasters/got-activations-llama3.1-70b-base.tabularfeature-extraction1K<n<10K0 likes370 downloads6mo agoHugging Face08HiTZ /Magpie-Llama-3.1-70B-Instruct-UnfilteredDataset generated using meta-llama/Llama-3.1-70B-Instruc with the MAGPIE codebase. The filtered dataset can be found here: HiTZ/Magpie-Llama-3.1-70B-Instruct-Filtered System prompts used General <|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nCutting Knowledge Date: December 2023\nToday Date: 26 Jul 2024\n\n<|eot_id|><|start_header_id|>user<|end_header_id|>\n\n Code <|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nYou are an AI… See the full description on the dataset page: https://huggingface.co/datasets/HiTZ/Magpie-Llama-3.1-70B-Instruct-Unfiltered.tabular1M<n<10M0 likes350 downloads1y agoHugging Face09penfever /meta-llama_Llama-3.1-70B-Instruct-jdgfct-Readabilitytext100K<n<1M0 likes308 downloads2y agoHugging Face10NousResearch /eval-Hermes-4-70B-nonreasoning hermes-70b-nonreasoning Evaluation Results Summary Benchmark Score Metric Samples Overlong rate aime24 0.095 math_pass@1:64_samples 64 99.4% aime25 0.073 math_pass@1:64_samples 64 98.2% arenahard 0.568 eval/overall_winrate 500 0.0% bbh_generative 0.805 extractive_match 1 100.0% creative-writing-v3 0.491 creative_writing_score 96 0.0% drop_generative_nous 0.784 drop_acc 1 100.0% eqbench3 0.739 eqbench_score 135 0.0% gpqa_diamond 0.333… See the full description on the dataset page: https://huggingface.co/datasets/NousResearch/eval-Hermes-4-70B-nonreasoning.tabular100K<n<1M3 likes305 downloads1y agoHugging Face11NousResearch /eval-Cogito-v2-preview-70B-reasoning cogito-thinking Evaluation Results Summary Benchmark Score Metric Samples Overlong rate aime24 0.322 math_pass@1:64_samples 64 35.2% aime25 0.221 math_pass@1:64_samples 64 33.3% arenahard 0.869 eval/overall_winrate 500 0.0% bbh_generative 0.893 extractive_match 1 2.9% creative-writing-v3 0.636 creative_writing_score 96 0.0% drop_generative_nous 0.860 drop_acc 1 0.8% eqbench3 0.657 eqbench_score 135 0.0% gpqa_diamond 0.591 gpqa_pass@1:8_samples8… See the full description on the dataset page: https://huggingface.co/datasets/NousResearch/eval-Cogito-v2-preview-70B-reasoning.tabular100K<n<1M2 likes298 downloads1y agoHugging Face12penfever /Nexusflow_Athene-70B-jdgfct-Completenesstext100K<n<1M0 likes278 downloads5mo agoHugging Face13NousResearch /eval-Cogito-v2-preview-70B-nonreasoning cogito-70b-nonthinking Evaluation Results Summary Benchmark Score Metric Samples Overlong rate aime24 0.122 math_pass@1:64_samples 64 100.0% aime25 0.060 math_pass@1:64_samples 64 100.0% arenahard 0.819 eval/overall_winrate 500 0.0% bbh_generative 0.876 extractive_match 1 100.0% creative-writing-v3 0.655 creative_writing_score 96 0.0% drop_generative_nous 0.841 drop_acc 1 100.0% eqbench3 0.681 eqbench_score 135 0.0% gpqa_diamond 0.528… See the full description on the dataset page: https://huggingface.co/datasets/NousResearch/eval-Cogito-v2-preview-70B-nonreasoning.tabular100K<n<1M2 likes277 downloads1y agoHugging Face14penfever /Nexusflow_Athene-70B-jdgfct-Harmlessnesstext100K<n<1M0 likes276 downloads2y agoHugging Face15OALL /details_sambanovasystems__SambaLingo-Arabic-Chat-70B Dataset Card for Evaluation run of sambanovasystems/SambaLingo-Arabic-Chat-70B Dataset automatically created during the evaluation run of model sambanovasystems/SambaLingo-Arabic-Chat-70B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_sambanovasystems__SambaLingo-Arabic-Chat-70B.tabular100K<n<1M0 likes243 downloads2y agoHugging Face16koyena /Magpie-Reasoning-V2-250K-CoT-Deepseek-R1-Llama-70B-formattedtext100K<n<1M0 likes220 downloads1y agoHugging Face17OALL /details_MaziyarPanahi__calme-2.3-llama3-70b Dataset Card for Evaluation run of MaziyarPanahi/calme-2.3-llama3-70b Dataset automatically created during the evaluation run of model MaziyarPanahi/calme-2.3-llama3-70b. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_MaziyarPanahi__calme-2.3-llama3-70b.tabular100K<n<1M0 likes206 downloads2y agoHugging Face18HiTZ /Magpie-Llama-3-70B-Instruct-UnfilteredDataset generated using meta-llama/Meta-Llama-3-70B-Instruct with the MAGPIE codebase. The filtered dataset can be found here: HiTZ/Magpie-Llama-3-70B-Instruct-Filtered System prompts used General <|begin_of_text|><|start_header_id|>user<|end_header_id|>\n\n Code <|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nYou are an AI assistant designed to provide helpful, step-by-step guidance on coding problems. The user will ask you a wide range of… See the full description on the dataset page: https://huggingface.co/datasets/HiTZ/Magpie-Llama-3-70B-Instruct-Unfiltered.tabular1M<n<10M0 likes203 downloads2y agoHugging Face19allenai /llama-3.1-tulu-3-70b-preference-mixture Llama 3.1 Tulu 3 70B Preference Mixture Note that this collection is licensed under ODC-BY-1.0 license; different licenses apply to subsets of the data. Some portions of the dataset are non-commercial. We present the mixture as a research artifact. This preference mixture used for DPO on our the Llama 3.1 Tulu 3 70B SFT checkpoint to obtain Llama 3.1 Tulu 3 70B DPO. This mix is made up from the following preference datasets:… See the full description on the dataset page: https://huggingface.co/datasets/allenai/llama-3.1-tulu-3-70b-preference-mixture.text100K<n<1M19 likes179 downloads2y agoHugging Face20NousResearch /eval-Hermes-4-70B-reasoning hermes-4-70b-reasoning-40k Evaluation Results Summary Benchmark Score Metric Samples Overlong rate aime24 0.735 math_pass@1:64_samples 64 8.4% aime25 0.674 math_pass@1:64_samples 64 9.6% arenahard 0.901 eval/overall_winrate 500 0.0% bbh_generative 0.878 extractive_match 1 4.8% creative-writing-v3 0.775 creative_writing_score 96 0.0% drop_generative_nous 0.850 drop_acc 1 1.4% eqbench3 0.847 eqbench_score 135 0.0% gpqa_diamond 0.661… See the full description on the dataset page: https://huggingface.co/datasets/NousResearch/eval-Hermes-4-70B-reasoning.tabular100K<n<1M5 likes179 downloads1y agoHugging Face21testcase-evaluate /all-Meta-Llama-3.1-70B-Instruct-AWQ-INT4text10M<n<100M0 likes176 downloads1y agoHugging Face22twinkle-ai /Llama-3.3-70B-Instruct-eval-logs-and-scorestabular100K<n<1M0 likes176 downloads7mo agoHugging Face23twinkle-ai /Llama-3-Taiwan-70B-Instruct-eval-logs-and-scorestabular100K<n<1M0 likes174 downloads7mo agoHugging Face24lightblue /reasoning-multilingual-R1-Llama-70B-train lightblue/reasoning-multilingual-R1-Llama-70B-train This is a multilingual reasoning dataset covering more than 30 languages. This dataset was made by: Sampling prompts from English datasets and translating them to various languages Generating responses to these prompts 8 times using deepseek-ai/DeepSeek-R1-Distill-Llama-70B Filtering out <think> sections with incorrect language, non-fluent language, and incorrect answers This dataset was then used to train a multilingual… See the full description on the dataset page: https://huggingface.co/datasets/lightblue/reasoning-multilingual-R1-Llama-70B-train.tabular1K<n<10K36 likes171 downloads2y agoHugging Face25OALL /details_airev-ai__Amal-70b-v5 Dataset Card for Evaluation run of airev-ai/Amal-70b-v5 Dataset automatically created during the evaluation run of model airev-ai/Amal-70b-v5. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_airev-ai__Amal-70b-v5.tabular100K<n<1M0 likes170 downloads2y agoHugging Face26hazyresearch /GPQA_with_Llama_3.1_70B_Instruct_v1 GPQA with Llama-3.1-70B-Instruct This dataset contains 646 graduate-level science questions from the GPQA benchmark with 100 candidate responses generated by Llama-3.1-70B-Instruct for each problem. Each response has been evaluated for correctness using a mixture of GPT-4o-mini and procedural Python code to robustly parse different answer formats, and scored by multiple reward models (scalar values) and LM judges (boolean verdicts). Dataset Structure Split: Single… See the full description on the dataset page: https://huggingface.co/datasets/hazyresearch/GPQA_with_Llama_3.1_70B_Instruct_v1.textn<1K0 likes163 downloads1y agoHugging Face27OALL /details_meta-llama__Llama-2-70b-hf Dataset Card for Evaluation run of meta-llama/Llama-2-70b-hf Dataset automatically created during the evaluation run of model meta-llama/Llama-2-70b-hf. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_meta-llama__Llama-2-70b-hf.tabular100K<n<1M0 likes162 downloads2y agoHugging Face28ReDiX /regolo-instruct-llama70B Regolo Instruct Llama-3.3-70B - Regolo.ai 🧠 Description This dataset was generated using Llama-3.3-70B, served via regolo.ai.The generation process was divided into two main stages: Translation of questions from open-source English-language datasets using Qwen2.5-7B Response generation through regolo Data { "messages": [ {"role": "system", "content": "<SYSTEM MESSAGE>"}, {"role": "user", "content":… See the full description on the dataset page: https://huggingface.co/datasets/ReDiX/regolo-instruct-llama70B.texttext-generation10K<n<100K3 likes156 downloads2y agoHugging Face29Magpie-Align /Magpie-Reasoning-V1-150K-CoT-Deepseek-R1-Llama-70B Project Web: https://magpie-align.github.io/ Arxiv Technical Report: https://arxiv.org/abs/2406.08464 Codes: https://github.com/magpie-align/magpie Abstract Click Here High-quality instruction data is critical for aligning large language models (LLMs). Although some models, such as Llama-3-Instruct, have open weights, their alignment data remain private, which hinders the democratization of AI. High human labor costs and a limited, predefined scope for prompting prevent… See the full description on the dataset page: https://huggingface.co/datasets/Magpie-Align/Magpie-Reasoning-V1-150K-CoT-Deepseek-R1-Llama-70B.text100K<n<1M18 likes138 downloads2y agoHugging Face30HiTZ /Magpie-Llama-3-70B-Instruct-FilteredDataset generated using meta-llama/Meta-Llama-3-70B-Instruct with the MAGPIE codebase. The unfiltered dataset can be found here: HiTZ/Magpie-Llama-3-70B-Instruct-Unfiltered Filter criteria def high_quality_filter(example): return ( example["input_quality"] in ["good", "excellent", "average"] and example["instruct_reward"] > -10 and not example["instruction"].endswith(":") and ( example["min_similar_conversation_id"] is None… See the full description on the dataset page: https://huggingface.co/datasets/HiTZ/Magpie-Llama-3-70B-Instruct-Filtered.tabular1M<n<10M0 likes137 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.