CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01SaylorTwift /details_meta-llama__Llama-3.1-8B-Instruct_private Dataset Card for Evaluation run of meta-llama/Llama-3.1-8B-Instruct Dataset automatically created during the evaluation run of model meta-llama/Llama-3.1-8B-Instruct. The dataset is composed of 78 configuration, each one corresponding to one of the evaluated task. The dataset has been created from 20 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/SaylorTwift/details_meta-llama__Llama-3.1-8B-Instruct_private.textn<1K0 likes4.5k downloads1y agoHugging Face02TAUR-Lab /Taur_CoT_Analysis_Project___meta-llama__Meta-Llama-3.1-8B-Instructtext10K<n<100K0 likes1.6k downloads2y agoHugging Face03Turbs /xprmt-llama-3.1-8b-instruct-multijail0 likes1.3k downloads5mo agoHugging Face04mksethi /llama-3.1-8b-Instruct_sae_repstext100K<n<1M0 likes565 downloads1y agoHugging Face05meta-llama /Llama-3.1-8B-Instruct-evalsgated Dataset Card for Llama-3.1-8B-Instruct Evaluation Result Details This dataset contains the Meta evaluation result details for Llama-3.1-8B-Instruct. The dataset has been created from 30 evaluation tasks. These tasks are human_eval, gorilla_api_bench__huggingface, mmlu_pro, infinite_bench, api_bank, human_eval_plus, ifeval__loose, mmlu__0_shot__cot, nih__multi_needle, multilingual_mmlu_de, mmlu, gsm8k, mgsm, multilingual_mmlu_fr, multilingual_mmlu_pt, math_hard… See the full description on the dataset page: https://huggingface.co/datasets/meta-llama/Llama-3.1-8B-Instruct-evals.text100K<n<1M34 likes438 downloads2y agoHugging Face06HiTZ /Magpie-Llama-3.1-8B-Instruct-UnfilteredDataset generated using meta-llama/Llama-3.1-8B-Instruc with the MAGPIE codebase. The filtered dataset can be found here: /HiTZ/Magpie-Llama-3.1-8B-Instruct-Filtered System prompts used General <|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nCutting Knowledge Date: December 2023\nToday Date: 26 Jul 2024\n\n<|eot_id|><|start_header_id|>user<|end_header_id|>\n\n Code <|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nYou are an AI… See the full description on the dataset page: https://huggingface.co/datasets/HiTZ/Magpie-Llama-3.1-8B-Instruct-Unfiltered.tabular1M<n<10M0 likes374 downloads1y agoHugging Face07penfever /meta-llama_Llama-3.1-8B-Instruct-jdgfct-Harmlessnesstext100K<n<1M0 likes348 downloads5mo agoHugging Face08HiTZ /Magpie-Llama-3.1-70B-Instruct-UnfilteredDataset generated using meta-llama/Llama-3.1-70B-Instruc with the MAGPIE codebase. The filtered dataset can be found here: HiTZ/Magpie-Llama-3.1-70B-Instruct-Filtered System prompts used General <|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nCutting Knowledge Date: December 2023\nToday Date: 26 Jul 2024\n\n<|eot_id|><|start_header_id|>user<|end_header_id|>\n\n Code <|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nYou are an AI… See the full description on the dataset page: https://huggingface.co/datasets/HiTZ/Magpie-Llama-3.1-70B-Instruct-Unfiltered.tabular1M<n<10M0 likes347 downloads1y agoHugging Face09Greenstar2 /demo-safety-overnight-Llama-3.1-8B-Instruct0 likes346 downloads16d agoHugging Face10penfever /meta-llama_Llama-3.1-8B-Instruct-jdgfct-Completenesstext100K<n<1M0 likes337 downloads5mo agoHugging Face11penfever /meta-llama_Llama-3.1-8B-Instruct-jdgfct-Readabilitytext100K<n<1M0 likes336 downloads6mo agoHugging Face12Amadeus99 /Llama-3.1-8B-Instruct-results0 likes336 downloads9mo agoHugging Face13penfever /meta-llama_Llama-3.1-70B-Instruct-jdgfct-Readabilitytext100K<n<1M0 likes308 downloads2y agoHugging Face14raymondzmc /tweet_topic_Llama-3.1-8B-Instruct_vocab_2000_lasttabular10K<n<100K0 likes272 downloads9mo agoHugging Face15raymondzmc /20_newsgroups_Llama-3.1-8B-Instruct_vocab_2000_lasttabular10K<n<100K0 likes228 downloads9mo agoHugging Face16raymondzmc /stackoverflow_Llama-3.1-8B-Instruct_vocab_2000_lasttabular10K<n<100K0 likes224 downloads9mo agoHugging Face17testcase-evaluate /all-Meta-Llama-3.1-70B-Instruct-AWQ-INT4text10M<n<100M0 likes216 downloads1y agoHugging Face18dongboklee /MMLU-Pro_Llama-3.1-8B-Instruct_gPRM_train MMLU-Pro_Llama-3.1-8B-Instruct_gPRM_train text10K<n<100K0 likes193 downloads3mo agoHugging Face19OALL /details_meta-llama__Llama-3.1-8B-Instruct_v2 Dataset Card for Evaluation run of meta-llama/Llama-3.1-8B-Instruct Dataset automatically created during the evaluation run of model meta-llama/Llama-3.1-8B-Instruct. The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_meta-llama__Llama-3.1-8B-Instruct_v2.text100K<n<1M0 likes192 downloads2y agoHugging Face20Greenstar2 /demo-safety-gate-features-Llama-3.1-8B-Instruct0 likes192 downloads21d agoHugging Face21andyrdt /maes-llama-3.1-8b-instruct0 likes190 downloads1y agoHugging Face22twinkle-ai /Llama-3.1-8B-Instruct-eval-logs-and-scorestabular100K<n<1M0 likes174 downloads7mo agoHugging Face23dongboklee /MMLU-Pro_Llama-3.1-8B-Instruct_gORM_train MMLU-Pro_Llama-3.1-8B-Instruct_gORM_train tabular100K<n<1M0 likes170 downloads3mo agoHugging Face24Turbs /xprmt-llama-3.1-8b-instruct-multijail-judge-evalimage0 likes168 downloads5mo agoHugging Face25twinkle-ai /Llama-3.1-Taiwan-8B-Instruct-eval-logs-and-scorestabular100K<n<1M0 likes168 downloads7mo agoHugging Face26hazyresearch /GPQA_with_Llama_3.1_70B_Instruct_v1 GPQA with Llama-3.1-70B-Instruct This dataset contains 646 graduate-level science questions from the GPQA benchmark with 100 candidate responses generated by Llama-3.1-70B-Instruct for each problem. Each response has been evaluated for correctness using a mixture of GPT-4o-mini and procedural Python code to robustly parse different answer formats, and scored by multiple reward models (scalar values) and LM judges (boolean verdicts). Dataset Structure Split: Single… See the full description on the dataset page: https://huggingface.co/datasets/hazyresearch/GPQA_with_Llama_3.1_70B_Instruct_v1.textn<1K0 likes163 downloads1y agoHugging Face27nishadsinghi /aime_solutions_llama_3.1_8B_instruct0 likes162 downloads2y agoHugging Face28OALL /details_Dampfinchen__Llama-3.1-8B-Ultra-Instruct Dataset Card for Evaluation run of Dampfinchen/Llama-3.1-8B-Ultra-Instruct Dataset automatically created during the evaluation run of model Dampfinchen/Llama-3.1-8B-Ultra-Instruct. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Dampfinchen__Llama-3.1-8B-Ultra-Instruct.tabular100K<n<1M0 likes144 downloads2y agoHugging Face29HINT-lab /Llama_3.1-8B-Instruct-Self-CalibrationThe official repository which contains the code and pre-trained models/datasets for our paper Efficient Test-Time Scaling via Self-Calibration. 🔥 Updates [2025-3-3]: We released our paper. [2025-2-25]: We released our codes, models and datasets. 🏴󠁶󠁵󠁭󠁡󠁰󠁿 Overview We propose an efficient test-time scaling method by using model confidence for dynamically sampling adjustment, since confidence can be seen as an intrinsic measure that directly reflects model… See the full description on the dataset page: https://huggingface.co/datasets/HINT-lab/Llama_3.1-8B-Instruct-Self-Calibration.tabularquestion-answering100K<n<1M0 likes137 downloads2y agoHugging Face30juiceb0xc0de /llama-3.1-8b-instruct-atlas llama-3.1-8b-instruct-atlas image1M<n<10M0 likes133 downloads27d agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.