CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01GulkoA /TinyStories-Llama-3.2-1B-cacheTinyStories dataset first layer activations by Llama-3.2-1B Useful for accelerated training and testing of sparse autoencoders hooked onto the first layer Context size: 128 tokens, batch size: 4 prompts 100k token version of this dataset: GulkoA/TinyStories-Llama-3.2-1B-cache-100k For tokenized dataset before activation caching, see GulkoA/TinyStories-tokenized-Llama-3.2 10K<n<100K0 likes290 downloads2y agoHugging Face02open-llm-leaderboard-old /details_freecs__Tiny-Llama-3-7b Dataset Card for Evaluation run of freecs/Tiny-Llama-3-7b Dataset automatically created during the evaluation run of model freecs/Tiny-Llama-3-7b on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_freecs__Tiny-Llama-3-7b.1 likes202 downloads3y agoHugging Face03Harvard-DCML /tis-subset-datasets-Llama-2-7b-hf Targeted Instruction Selection Subsets (Llama-2-7b-hf) This repository contains pre-computed instruction training subsets selected from a large candidate pool for targeted instruction fine-tuning, as presented in the paper A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't). Paper: https://huggingface.co/papers/2602.14696 GitHub Repository: https://github.com/dcml-lab/targeted-instruction-selection Description Instruction… See the full description on the dataset page: https://huggingface.co/datasets/Harvard-DCML/tis-subset-datasets-Llama-2-7b-hf.texttext-generation100K<n<1M0 likes177 downloads7mo agoHugging Face04OALL /details_gaverfraxz__Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES Dataset Card for Evaluation run of gaverfraxz/Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES Dataset automatically created during the evaluation run of model gaverfraxz/Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_gaverfraxz__Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES.tabular100K<n<1M0 likes130 downloads2y agoHugging Face05GulkoA /TinyStories-tokenized-Llama-3.2-1024-contextTinyStories dataset tokenized with Llama-3.2 Useful for accelerated training and testing of sparse autoencoders Context window: 1024, not shuffled 1K<n<10K0 likes130 downloads2y agoHugging Face06GulkoA /TinyStories-tokenized-Llama-3.2TinyStories dataset tokenized with Llama-3.2 Useful for accelerated training and testing of sparse autoencoders Context window: 128, not shuffled For first layer activations cache with Llama-3.2-1B, see GulkoA/TinyStories-Llama-3.2-1B-cache text-generation1M<n<10M1 likes125 downloads2y agoHugging Face07GulkoA /TinyStories-Llama-3.2-1B-cache-layer-5batch_size: 1024 prompts training_tokens: 1,000,000 hook_layer: 5 hook_name: blocks.5.hook_mlp_out 1K<n<10K1 likes101 downloads1y agoHugging Face08open-llm-leaderboard-old /details_BEE-spoke-data__smol_llama-81M-tied Dataset Card for Evaluation run of BEE-spoke-data/smol_llama-81M-tied Dataset Summary Dataset automatically created during the evaluation run of model BEE-spoke-data/smol_llama-81M-tied on the Open LLM Leaderboard. The dataset is composed of 64 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_BEE-spoke-data__smol_llama-81M-tied.0 likes68 downloads3y agoHugging Face09princenandal /tiny-llama-hint-gentext10K<n<100K0 likes61 downloads2y agoHugging Face10open-llm-leaderboard /vhab10__Llama-3.2-Instruct-3B-TIES-detailsgated Dataset Card for Evaluation run of vhab10/Llama-3.2-Instruct-3B-TIES Dataset automatically created during the evaluation run of model vhab10/Llama-3.2-Instruct-3B-TIES The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/vhab10__Llama-3.2-Instruct-3B-TIES-details.tabular10K<n<100K0 likes48 downloads2y agoHugging Face11Harvard-DCML /tis-dolci-subset-datasets-Llama-3.2-3Btext100K<n<1M0 likes48 downloads4mo agoHugging Face12open-llm-leaderboard /khoantap__llama-evolve-ties-best-merge-detailsgated Dataset Card for Evaluation run of khoantap/llama-evolve-ties-best-merge Dataset automatically created during the evaluation run of model khoantap/llama-evolve-ties-best-merge The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/khoantap__llama-evolve-ties-best-merge-details.tabular10K<n<100K0 likes39 downloads2y agoHugging Face13tim-lawson /open-web-math_TinyLlama_v1.1_KLdiv_Llama-2-7b-hf_TinyLlama_v1.10 likes39 downloads1y agoHugging Face14Harvard-DCML /tis-quantile-datasets-Llama-3.2-3Btext10K<n<100K0 likes36 downloads7mo agoHugging Face15JasonYan777 /PersonaSignal-LeakageCheck-Locale-And-Time-Zone-Meta-Llama-3.1-8B-Instruct-Turbotabularn<1K0 likes33 downloads10mo agoHugging Face16qazisaad /llama_2_optimized_product_titles-esci-test-sft Dataset Card for "llama_2_optimized_product_titles-esci-test-sft" More Information needed tabular10K<n<100K0 likes30 downloads3y agoHugging Face17Harvard-DCML /tis-dolci-subset-datasets-Llama-2-7b-hftext100K<n<1M0 likes30 downloads4mo agoHugging Face18qazisaad /llama_2_product_titles-esci_train-temp-pos Dataset Card for "llama_2_product_titles-esci_train-temp-pos" More Information needed tabular1K<n<10K0 likes28 downloads3y agoHugging Face19qazisaad /llama-2-optimized-product-titles-esci-4-7-temp Dataset Card for "llama-2-optimized-product-titles-esci-4-7-temp" More Information needed tabular1K<n<10K0 likes28 downloads3y agoHugging Face20open-llm-leaderboard-old /details_Gryphe__Tiamat-8b-1.2-Llama-3-DPO0 likes27 downloads2y agoHugging Face21Harvard-DCML /tis-quantile-datasets-Llama-2-7b-hftext10K<n<100K0 likes24 downloads7mo agoHugging Face22qazisaad /llama_2-product-titles-esci-test-temp Dataset Card for "llama_2-product-titles-esci-test-temp" More Information needed tabular1K<n<10K0 likes22 downloads3y agoHugging Face23Harvard-DCML /tis-subset-datasets-Llama-3.2-3Btext100K<n<1M0 likes20 downloads7mo agoHugging Face24qazi-ali /llama_2-optimized-titles-esci-sft-test Dataset Card for "llama_2-optimized-titles-esci-sft-test" More Information needed tabular1K<n<10K0 likes19 downloads3y agoHugging Face25qazisaad /llama-2-optimized-product-titles-esci-test-sft-temp Dataset Card for "llama-2-optimized-product-titles-esci-test-sft-temp" More Information needed tabular1K<n<10K0 likes19 downloads3y agoHugging Face26qazisaad /llama_2-product-titles-esci-test-sft-temp Dataset Card for "llama_2-product-titles-esci-test-sft-temp" More Information needed tabular10K<n<100K0 likes18 downloads3y agoHugging Face27qazi-ali /llama_2-product-titles-esci-sft-train Dataset Card for "llama_2-product-titles-esci-sft-train" More Information needed tabular1K<n<10K0 likes15 downloads3y agoHugging Face28MadhuriD /New_Tiny_llama_Madhuritextn<1K0 likes15 downloads2y agoHugging Face29yaya-sy /rp_test_tiny_llamatext10K<n<100K0 likes15 downloads2y agoHugging Face30GulkoA /TinyStories-Llama-3.2-1B-cache-100kTinyStories dataset first layer activations by Llama-3.2-1B Useful for accelerated training and testing of sparse autoencoders hooked onto the first layer Context size: 128 tokens, batch size: 4 prompts, limited to 100k input tokens For tokenized dataset before activation caching, see GulkoA/TinyStories-tokenized-Llama-3.2 text-generation1K<n<10K0 likes15 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.