CoolFace
20 results

llama-ti

GulkoA /TinyStories-Llama-3.2-1B-cacheTinyStories dataset first layer activations by Llama-3.2-1B Useful for accelerated training and testing of sparse autoencoders hooked onto the first layer Context size: 128 tokens, batch size: 4 prompts 100k token version of this dataset: GulkoA/TinyStories-Llama-3.2-1B-cache-100k For tokenized dataset before activation caching, see GulkoA/TinyStories-tokenized-Llama-3.2 10K<n<100K0 likes229 downloads2y agoHugging Faceopen-llm-leaderboard-old /details_freecs__Tiny-Llama-3-7b Dataset Card for Evaluation run of freecs/Tiny-Llama-3-7b Dataset automatically created during the evaluation run of model freecs/Tiny-Llama-3-7b on the Open LLM Leaderboard. The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_freecs__Tiny-Llama-3-7b.1 likes200 downloads3y agoHugging FaceHarvard-DCML /tis-subset-datasets-Llama-2-7b-hf Targeted Instruction Selection Subsets (Llama-2-7b-hf) This repository contains pre-computed instruction training subsets selected from a large candidate pool for targeted instruction fine-tuning, as presented in the paper A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't). Paper: https://huggingface.co/papers/2602.14696 GitHub Repository: https://github.com/dcml-lab/targeted-instruction-selection Description Instruction… See the full description on the dataset page: https://huggingface.co/datasets/Harvard-DCML/tis-subset-datasets-Llama-2-7b-hf.texttext-generation100K<n<1M0 likes181 downloads7mo agoHugging FaceGulkoA /TinyStories-tokenized-Llama-3.2-1024-contextTinyStories dataset tokenized with Llama-3.2 Useful for accelerated training and testing of sparse autoencoders Context window: 1024, not shuffled 1K<n<10K0 likes131 downloads2y agoHugging FaceOALL /details_gaverfraxz__Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES Dataset Card for Evaluation run of gaverfraxz/Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES Dataset automatically created during the evaluation run of model gaverfraxz/Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES. The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task. The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_gaverfraxz__Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES.tabular100K<n<1M0 likes128 downloads2y agoHugging FaceGulkoA /TinyStories-tokenized-Llama-3.2TinyStories dataset tokenized with Llama-3.2 Useful for accelerated training and testing of sparse autoencoders Context window: 128, not shuffled For first layer activations cache with Llama-3.2-1B, see GulkoA/TinyStories-Llama-3.2-1B-cache text-generation1M<n<10M1 likes119 downloads2y agoHugging Face