llama-ti
Datasets
All datasets matching “llama-ti”TinyStories-Llama-3.2-1B-cacheTinyStories dataset first layer activations by Llama-3.2-1B
Useful for accelerated training and testing of sparse autoencoders hooked onto the first layer
Context size: 128 tokens, batch size: 4 prompts
100k token version of this dataset: GulkoA/TinyStories-Llama-3.2-1B-cache-100k
For tokenized dataset before activation caching, see GulkoA/TinyStories-tokenized-Llama-3.2
details_freecs__Tiny-Llama-3-7b
Dataset Card for Evaluation run of freecs/Tiny-Llama-3-7b
Dataset automatically created during the evaluation run of model freecs/Tiny-Llama-3-7b on the Open LLM Leaderboard.
The dataset is composed of 63 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard-old/details_freecs__Tiny-Llama-3-7b.tis-subset-datasets-Llama-2-7b-hf
Targeted Instruction Selection Subsets (Llama-2-7b-hf)
This repository contains pre-computed instruction training subsets selected from a large candidate pool for targeted instruction fine-tuning, as presented in the paper A Critical Look at Targeted Instruction Selection: Disentangling What Matters (and What Doesn't).
Paper: https://huggingface.co/papers/2602.14696
GitHub Repository: https://github.com/dcml-lab/targeted-instruction-selection
Description
Instruction… See the full description on the dataset page: https://huggingface.co/datasets/Harvard-DCML/tis-subset-datasets-Llama-2-7b-hf.TinyStories-tokenized-Llama-3.2-1024-contextTinyStories dataset tokenized with Llama-3.2
Useful for accelerated training and testing of sparse autoencoders
Context window: 1024, not shuffled
details_gaverfraxz__Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES
Dataset Card for Evaluation run of gaverfraxz/Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES
Dataset automatically created during the evaluation run of model gaverfraxz/Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES.
The dataset is composed of 136 configuration, each one coresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The… See the full description on the dataset page: https://huggingface.co/datasets/OALL/details_gaverfraxz__Meta-Llama-3.1-8B-Instruct-HalfAbliterated-TIES.TinyStories-tokenized-Llama-3.2TinyStories dataset tokenized with Llama-3.2
Useful for accelerated training and testing of sparse autoencoders
Context window: 128, not shuffled
For first layer activations cache with Llama-3.2-1B, see GulkoA/TinyStories-Llama-3.2-1B-cache
