CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01scaleinvariant /paired-llama-3.2-1b-embeddings-lmsys-chat-1m Paired Llama 3.2 1B Token Embeddings (LMSYS-Chat-1M) This dataset contains paired activations corresponding to single token locations extracted from Meta's Llama 3.2 1B Instruct on conversations from LMSYS-Chat-1M. Embeddings are provided for layers 5 through 14, which capture the most interesting intermediate representations. This dataset was built to study things like: Learning different basis for activations at a given layer Studying if there are cases where position encodes… See the full description on the dataset page: https://huggingface.co/datasets/scaleinvariant/paired-llama-3.2-1b-embeddings-lmsys-chat-1m.tabularfeature-extraction100M<n<1B3 likes6.5k downloads7mo agoHugging Face02juiceb0xc0de /llama-3.2-1b-atlas llama-3.2-1b-atlas image100K<n<1M1 likes1.9k downloads26d agoHugging Face03toksuitebackup /meta-llama-Llama-3.2-1B-toksuite-detokenizedTraining data of the model detokenized in the exact order seen by the model. The training data is partitioned into 8 chunks (chunk-0 through chunk-7), based on the GPU rank that generated the data. Each chunk contains detokenized text files in JSON Lines format (.jsonl). tabular10M<n<100M0 likes1.2k downloads10mo agoHugging Face04ENSEONG /full-math-private-n256-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes1.2k downloads5mo agoHugging Face05ENSEONG /preprocessed-full-math-private-n256-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes967 downloads5mo agoHugging Face06jan-hq /instruction-convert-audio-whispervq-llama3.2-compresstext1M<n<10M0 likes658 downloads2y agoHugging Face07sehyun734 /longhealth-llama-3.2-3btext10K<n<100K1 likes436 downloads5d agoHugging Face08jan-hq /instruction-convert-audio-whispervq-llama3.2text1M<n<10M0 likes382 downloads2y agoHugging Face09scaleinvariant /llama-3.2-1b-instruct-lmsys-chat-1m-activations Llama 3.2 1B Instruct Activations (LMSYS-Chat-1M) This dataset contains whole-model residual stream activations extracted from Meta's Llama 3.2 1B Instruct on conversations from LMSYS-Chat-1M. Each row stores the complete residual stream across all 16 transformer layers for a single prompt — both the full-sequence activations and the final-token activations. Note: This is a subset, 8% (from 2 workers of 25) of the full dataset. The complete dataset was ~25 TB and huggingface only… See the full description on the dataset page: https://huggingface.co/datasets/scaleinvariant/llama-3.2-1b-instruct-lmsys-chat-1m-activations.textfeature-extraction10K<n<100K0 likes382 downloads6mo agoHugging Face10nishadsinghi /MATH_train_llama3.2-3b-instructtext1K<n<10K0 likes377 downloads2y agoHugging Face11jan-hq /instruction-convert-audio-whispervq-llama3.2-deduptext1M<n<10M0 likes367 downloads2y agoHugging Face12ENSEONG /preprocessed-full-math-private-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes304 downloads6mo agoHugging Face13juiceb0xc0de /llama-3.2-3b-atlas llama-3.2-3b-atlas image1M<n<10M0 likes294 downloads26d agoHugging Face14ENSEONG /full-math-private-Llama-3.2-3B-Instruct-bontabular100K<n<1M0 likes264 downloads6mo agoHugging Face15juiceb0xc0de /llama3.2-3b-instruct-atlas juiceb0xc0de/llama3.2-3b-instruct-atlas A brain atlas for meta-llama/Llama-3.2-3B-Instruct, the 3B instruction-tuned member of the Llama 3.2 family. This is not a chat dataset or a benchmark - it is an internal-mechanics map built by running activations through a corpus of prompts and scoring what each layer, component, head, and feature direction is doing. If you want to know where an instruction-tuned model keeps its register machinery, which directions survive a causal test… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/llama3.2-3b-instruct-atlas.imagefeature-extraction1M<n<10M0 likes249 downloads26d agoHugging Face16evinsi /fineweb-edu-Llama-3.2-Instruct-Shuffledtext1M<n<10M0 likes228 downloads2y agoHugging Face17crystal-ai /chat-compilation-benchmark-5x-Llama-3.2-Instruct-Shuffledtext1M<n<10M0 likes207 downloads1y agoHugging Face18OALL /details_meta-llama__Llama-3.2-3B-Instruct_v2 Dataset Card for Evaluation run of meta-llama/Llama-3.2-3B-Instruct Dataset automatically created during the evaluation run of model meta-llama/Llama-3.2-3B-Instruct. The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_meta-llama__Llama-3.2-3B-Instruct_v2.text100K<n<1M0 likes193 downloads1y agoHugging Face19juiceb0xc0de /llama3.2-1b-instruct-atlas juiceb0xc0de/llama3.2-1b-instruct-atlas A brain atlas for meta-llama/Llama-3.2-1B-Instruct, the 1B instruction-tuned member of the Llama 3.2 family. This is not a chat dataset or a benchmark - it is an internal-mechanics map built by running activations through a corpus of prompts and scoring what each layer, component, head, and feature direction is doing. If you want to know where a small instruction-tuned model keeps its register machinery, which directions survive a causal… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/llama3.2-1b-instruct-atlas.imagefeature-extraction1M<n<10M0 likes190 downloads26d agoHugging Face20HuggingFaceH4 /Llama-3.2-1B-Instruct-best-of-N-completionstabular1K<n<10K1 likes180 downloads2y agoHugging Face21twinkle-ai /llama-3.2-3B-f1-instruct-eval-logs-and-scorestabular100K<n<1M0 likes161 downloads7mo agoHugging Face22twinkle-ai /Llama-3.2-3B-Instruct-eval-logs-and-scorestabular100K<n<1M0 likes160 downloads7mo agoHugging Face23yosefw /magpie-llama-3.2-1b-instructtext100K<n<1M0 likes159 downloads11d agoHugging Face24SaylorTwift /RULER-16384-llama-3.2-tokenizertabular1K<n<10K0 likes145 downloads1y agoHugging Face25OALL /details_meta-llama__Llama-3.2-1B_v2 Dataset Card for Evaluation run of meta-llama/Llama-3.2-1B Dataset automatically created during the evaluation run of model meta-llama/Llama-3.2-1B. The dataset is composed of 116 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest results. An additional… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_meta-llama__Llama-3.2-1B_v2.text100K<n<1M0 likes139 downloads1y agoHugging Face26JakeOh /star_plus-llama-3.2-1b-gsm8k-step-3text100K<n<1M0 likes134 downloads2y agoHugging Face27sibasmarakp /Llama-3.2-1B-Instruct-uPRM-T80-adapters-dvts-completionstabular1K<n<10K0 likes122 downloads8mo agoHugging Face28hubnemo /mtp-selfdata-llama3.2-3b-finewikitext10K<n<100K0 likes122 downloads16d agoHugging Face29HuggingFaceH4 /Llama-3.2-1B-Instruct-beam-search-completionstabular10K<n<100K1 likes120 downloads2y agoHugging Face30SaylorTwift /RULER-4096-llama-3.2-tokenizertabular1K<n<10K0 likes120 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.