CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01OALL /details_grimjim__Llama-3-Instruct-8B-SimPO-SPPO-Iter3-merge Dataset Card for Evaluation run of grimjim/Llama-3-Instruct-8B-SimPO-SPPO-Iter3-merge Dataset automatically created during the evaluation run of model grimjim/Llama-3-Instruct-8B-SimPO-SPPO-Iter3-merge. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_grimjim__Llama-3-Instruct-8B-SimPO-SPPO-Iter3-merge.tabular100K<n<1M0 likes2k downloads2y agoHugging Face02sammyliu /qwen3-8b-activations-l20-l36 Qwen3 8B Activations for Layers 20 and 36 This dataset contains assistant-token residual activations harvested from Qwen/Qwen3-8B over 980000 training conversations from lmsys/lmsys-chat-1m. We only generated for Layer 20 and 36 because each one costs 2TB and we simply cannot afford to store more :) You can use this dataset to train SAEs, linear probes, other mech interp models etc, for Qwen3 8B. We picked Qwen3 8B because this is a small part of a larger experiment to use feature… See the full description on the dataset page: https://huggingface.co/datasets/sammyliu/qwen3-8b-activations-l20-l36.tabular100M<n<1B0 likes1.7k downloads6mo agoHugging Face03scaleinvariant /sae-activations-llama-3.1-8b-layer19-lmsys-chat-1m SAE Feature Activations — Llama 3.1 8B Instruct, Layer 19 (LMSYS-Chat-1M) This dataset contains Sparse Autoencoder (SAE) feature activations extracted from layer 19 of Meta's Llama 3.1 8B Instruct on conversations from LMSYS-Chat-1M. It also has natural language explainations of features generated by GPT OSS 120B. See subset 4 for details. The SAE used is Goodfire/Llama-3.1-8B-Instruct-SAE-l19, which decomposes layer-19 residual stream activations into interpretable sparse features.… See the full description on the dataset page: https://huggingface.co/datasets/scaleinvariant/sae-activations-llama-3.1-8b-layer19-lmsys-chat-1m.tabularfeature-extraction100M<n<1B0 likes1.2k downloads6mo agoHugging Face04OALL /details_princeton-nlp__Llama-3-8B-ProLong-512k-Instruct Dataset Card for Evaluation run of princeton-nlp/Llama-3-8B-ProLong-512k-Instruct Dataset automatically created during the evaluation run of model princeton-nlp/Llama-3-8B-ProLong-512k-Instruct. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_princeton-nlp__Llama-3-8B-ProLong-512k-Instruct.tabular100K<n<1M0 likes1.1k downloads2y agoHugging Face05toksuitebackup /aya-expanse-8b-toksuite-detokenizedTraining data of the model detokenized in the exact order seen by the model. The training data is partitioned into 8 chunks (chunk-0 through chunk-7), based on the GPU rank that generated the data. Each chunk contains detokenized text files in JSON Lines format (.jsonl). tabular10M<n<100M0 likes854 downloads10mo agoHugging Face06Harsh01012 /hubble-8b-unlearning-resultsimage1K<n<10K1 likes814 downloads2mo agoHugging Face07argo11 /0399-tv-valid-clean-sft-tokenized-llmjp4-8btabular1M<n<10M0 likes796 downloads3mo agoHugging Face08nyu-dice-lab /lm-eval-results-princeton-nlp-Llama-3-Base-8B-SFT-RDPO-private Dataset Card for Evaluation run of princeton-nlp/Llama-3-Base-8B-SFT-RDPO Dataset automatically created during the evaluation run of model princeton-nlp/Llama-3-Base-8B-SFT-RDPO The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-princeton-nlp-Llama-3-Base-8B-SFT-RDPO-private.tabular100K<n<1M0 likes697 downloads2y agoHugging Face09sunshineNew /rh_qwen3_8b_prompted_v2_completions TRL Completion logs This dataset contains the completions generated during training using trl. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the completion generated by the model <reward_function_name>: the reward(s) assigned to the… See the full description on the dataset page: https://huggingface.co/datasets/sunshineNew/rh_qwen3_8b_prompted_v2_completions.tabular1K<n<10K0 likes595 downloads13d agoHugging Face10RLAIF /numina-math-llama-3.1-8b-bon-meta-cottabular100K<n<1M0 likes543 downloads2y agoHugging Face11sunshineNew /rh_qwen3_8b_sdf_completions TRL Completion logs This dataset contains the completions generated during training using trl. The completions are stored in parquet files, and each file contains the completions for a single step of training (depending on the logging_steps argument). Each file contains the following columns: step: the step of training prompt: the prompt used to generate the completion completion: the completion generated by the model <reward_function_name>: the reward(s) assigned to the… See the full description on the dataset page: https://huggingface.co/datasets/sunshineNew/rh_qwen3_8b_sdf_completions.tabular1K<n<10K0 likes500 downloads15d agoHugging Face12OALL /details_meta-llama__Meta-Llama-3-8B-Instruct Dataset Card for Evaluation run of meta-llama/Meta-Llama-3-8B-Instruct Dataset automatically created during the evaluation run of model meta-llama/Meta-Llama-3-8B-Instruct. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_meta-llama__Meta-Llama-3-8B-Instruct.tabular100K<n<1M0 likes442 downloads2y agoHugging Face13nyu-dice-lab /lm-eval-results-hkust-nlp-dart-math-llama3-8b-prop2diff-private Dataset Card for Evaluation run of hkust-nlp/dart-math-llama3-8b-prop2diff Dataset automatically created during the evaluation run of model hkust-nlp/dart-math-llama3-8b-prop2diff The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 4 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-hkust-nlp-dart-math-llama3-8b-prop2diff-private.tabular100K<n<1M0 likes416 downloads2y agoHugging Face14juiceb0xc0de /qwen3-8b-base-atlas-SAE Qwen3-8B-Base Feature Atlas A single queryable SQLite database (atlas.sqlite, ~570 MB) that maps the internals of Qwen/Qwen3-8B-Base — every weight channel and every sparse-autoencoder feature scored for what it selects for, across a register-diverse corpus of 4,946 prompts. It is not a text dataset. There are no training rows. It is an index of model internals — the kind of thing you query to find "which channels in layer 23 discriminate compliance from authentic-personality… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/qwen3-8b-base-atlas-SAE.tabular1M<n<10M1 likes389 downloads7d agoHugging Face15qqq12311 /qwen3-8b-residualstabularn<1K0 likes364 downloads2mo agoHugging Face16fffoivos /apertus-8b-greek-cpt-modern-greek-train Exact Modern-Greek training content for Apertus 8B Greek CPT This is the public Modern-Greek, train-only document snapshot selected for the full 8B D0 continued-pretraining run. It preserves the upstream v2 schema and metadata; text is reproduced as its exact training-time Apertus-parity PII-masked value. Selection is reconstructed from immutable post-mask training catalogs and content hashes. It contains no replay payload. Exact selected content HPLT Modern… See the full description on the dataset page: https://huggingface.co/datasets/fffoivos/apertus-8b-greek-cpt-modern-greek-train.tabular10M<n<100M0 likes342 downloads1mo agoHugging Face17weqweasdas /new_8b_self_corr_standardtabular1M<n<10M0 likes316 downloads2y agoHugging Face18HiTZ /Magpie-Llama-3.1-8B-Instruct-UnfilteredDataset generated using meta-llama/Llama-3.1-8B-Instruc with the MAGPIE codebase. The filtered dataset can be found here: /HiTZ/Magpie-Llama-3.1-8B-Instruct-Filtered System prompts used General <|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nCutting Knowledge Date: December 2023\nToday Date: 26 Jul 2024\n\n<|eot_id|><|start_header_id|>user<|end_header_id|>\n\n Code <|begin_of_text|><|start_header_id|>system<|end_header_id|>\n\nYou are an AI… See the full description on the dataset page: https://huggingface.co/datasets/HiTZ/Magpie-Llama-3.1-8B-Instruct-Unfiltered.tabular1M<n<10M0 likes308 downloads1y agoHugging Face19prince-canuma /fineweb-CC-MAIN-2024-10-8B-entabular10M<n<100M0 likes297 downloads2y agoHugging Face20toksuite /Qwen-Qwen3-8B-toksuite-detokenizedTraining data of the model detokenized in the exact order seen by the model. The training data is partitioned into 8 chunks (chunk-0 through chunk-7), based on the GPU rank that generated the data. Each chunk contains detokenized text files in JSON Lines format (.jsonl). tabular10M<n<100M0 likes296 downloads9mo agoHugging Face21OALL /details_Orenguteng__Llama-3.1-8B-Lexi-Uncensored-V2 Dataset Card for Evaluation run of Orenguteng/Llama-3.1-8B-Lexi-Uncensored-V2 Dataset automatically created during the evaluation run of model Orenguteng/Llama-3.1-8B-Lexi-Uncensored-V2. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_Orenguteng__Llama-3.1-8B-Lexi-Uncensored-V2.tabular100K<n<1M0 likes295 downloads2y agoHugging Face22nyu-dice-lab /lm-eval-results-allenai-llama-3-tulu-2-dpo-8b-private Dataset Card for Evaluation run of allenai/llama-3-tulu-2-dpo-8b Dataset automatically created during the evaluation run of model allenai/llama-3-tulu-2-dpo-8b The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-allenai-llama-3-tulu-2-dpo-8b-private.tabular100K<n<1M0 likes259 downloads2y agoHugging Face23unlearning-cleanslate /generations-llama-3_1-8b-rmu-baselinetabular10K<n<100K0 likes259 downloads5mo agoHugging Face24nyu-dice-lab /lm-eval-results-penfever-Llama-3-8B-tulu-human-v2-private Dataset Card for Evaluation run of penfever/Llama-3-8B-tulu-human-v2 Dataset automatically created during the evaluation run of model penfever/Llama-3-8B-tulu-human-v2 The dataset is composed of 62 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 7 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/nyu-dice-lab/lm-eval-results-penfever-Llama-3-8B-tulu-human-v2-private.tabular100K<n<1M0 likes254 downloads2y agoHugging Face25juiceb0xc0de /qwen3-8b-atlas juiceb0xc0de/qwen3-8b-atlas A brain atlas for Qwen/Qwen3-8B, the 8B dense member of the Qwen3 family. This is not a chat dataset or a benchmark. It is an internal-mechanics map built by running activations through a corpus of prompts and scoring what each layer, component, head, and feature direction is doing. If you want to know what a model with no idle capacity looks like from the inside, where register information lives in a well-trained dense stack, or why this particular… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/qwen3-8b-atlas.imagefeature-extraction1M<n<10M0 likes252 downloads25d agoHugging Face26anime-sh /generations-llama-3_1-8b-rmu-bm25-10b-rebuttaltabular10K<n<100K0 likes252 downloads2mo agoHugging Face27toksuitebackup /Qwen-Qwen3-8B-toksuite-detokenizedTraining data of the model detokenized in the exact order seen by the model. The training data is partitioned into 8 chunks (chunk-0 through chunk-7), based on the GPU rank that generated the data. Each chunk contains detokenized text files in JSON Lines format (.jsonl). tabular10M<n<100M0 likes247 downloads9mo agoHugging Face28OALL /details_terrycraddock__Reflection-Llama-3.1-8B Dataset Card for Evaluation run of terrycraddock/Reflection-Llama-3.1-8B Dataset automatically created during the evaluation run of model terrycraddock/Reflection-Llama-3.1-8B. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_terrycraddock__Reflection-Llama-3.1-8B.tabular100K<n<1M0 likes243 downloads2y agoHugging Face29OALL /details_EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 Dataset Card for Evaluation run of EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200 Dataset automatically created during the evaluation run of model EpistemeAI/Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train"… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_EpistemeAI__Fireball-Alpaca-Llama-3.1-8B-Philos-DPO-200.tabular100K<n<1M0 likes226 downloads2y agoHugging Face30anime-sh /generations-llama-3_1-8b-undial-bm25-6t-rebuttaltabular10K<n<100K0 likes222 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.