datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
paired-llama-3.2-1b-embeddings-lmsys-chat-1m
Paired Llama 3.2 1B Token Embeddings (LMSYS-Chat-1M)
This dataset contains paired activations corresponding to single token locations extracted from Meta's Llama 3.2 1B Instruct on conversations from LMSYS-Chat-1M.
Embeddings are provided for layers 5 through 14, which capture the most interesting intermediate representations.
This dataset was built to study things like:
Learning different basis for activations at a given layer
Studying if there are cases where position encodes… See the full description on the dataset page: https://huggingface.co/datasets/scaleinvariant/paired-llama-3.2-1b-embeddings-lmsys-chat-1m.llama-3.2-1b-atlas
llama-3.2-1b-atlas
full-math-private-n256-Llama-3.2-3B-Instruct-bonpreprocessed-full-math-private-n256-Llama-3.2-3B-Instruct-bonmeta-llama-Llama-3.2-1B-toksuite-detokenizedTraining data of the model detokenized in the exact order seen by the model.
The training data is partitioned into 8 chunks (chunk-0 through chunk-7), based on the GPU rank that generated the data. Each chunk contains detokenized text files in JSON Lines format (.jsonl).
preprocessed-full-math-private-Llama-3.2-3B-Instruct-bonfull-math-private-Llama-3.2-3B-Instruct-bonllama3.2-3b-instruct-atlas
juiceb0xc0de/llama3.2-3b-instruct-atlas
A brain atlas for meta-llama/Llama-3.2-3B-Instruct, the 3B instruction-tuned member of the Llama 3.2 family. This is not a chat dataset or a benchmark - it is an internal-mechanics map built by running activations through a corpus of prompts and scoring what each layer, component, head, and feature direction is doing.
If you want to know where an instruction-tuned model keeps its register machinery, which directions survive a causal test… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/llama3.2-3b-instruct-atlas.llama-3.2-3b-atlas
llama-3.2-3b-atlas
llama3.2-1b-instruct-atlas
juiceb0xc0de/llama3.2-1b-instruct-atlas
A brain atlas for meta-llama/Llama-3.2-1B-Instruct, the 1B instruction-tuned member of the Llama 3.2 family. This is not a chat dataset or a benchmark - it is an internal-mechanics map built by running activations through a corpus of prompts and scoring what each layer, component, head, and feature direction is doing.
If you want to know where a small instruction-tuned model keeps its register machinery, which directions survive a causal… See the full description on the dataset page: https://huggingface.co/datasets/juiceb0xc0de/llama3.2-1b-instruct-atlas.llama-3.2-3B-f1-instruct-eval-logs-and-scoresLlama-3.2-3B-Instruct-eval-logs-and-scoresLlama-3.2-1B-Instruct-best-of-N-completionsRULER-16384-llama-3.2-tokenizerRULER-4096-llama-3.2-tokenizerLlama-3.2-1B-Instruct-beam-search-completionsLlama-3.2-1B-Instruct-uPRM-T80-adapters-dvts-completionsLlama-3.2-3B-Instruct-beam-search-completionsLlama-3.2-1B-Instruct-DVTS-completionsLlama-3.2-3B-Instruct-best-of-N-completionsRULER-131072-llama-3.2-tokenizerreward-bench-Llama-3.2-1B-yes-noreward-bench-Llama-3.2-3B-yes-noWhole-Data-Llama-3.2-3B-Instruct-20_armo_tokenizedLlama-3.2-1B-Instruct-uPRM-70B-T80-olympiadbench-best_of_n-completionsakhadangi__Llama3.2.1B.0.01-First-details
Dataset Card for Evaluation run of akhadangi/Llama3.2.1B.0.01-First
Dataset automatically created during the evaluation run of model akhadangi/Llama3.2.1B.0.01-First
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 1 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/akhadangi__Llama3.2.1B.0.01-First-details.RULER-32768-llama-3.2-tokenizerNousResearch__Hermes-3-Llama-3.2-3B-details
Dataset Card for Evaluation run of NousResearch/Hermes-3-Llama-3.2-3B
Dataset automatically created during the evaluation run of model NousResearch/Hermes-3-Llama-3.2-3B
The dataset is composed of 38 configuration(s), each one corresponding to one of the evaluated task.
The dataset has been created from 2 run(s). Each run can be found as a specific split in each configuration, the split being named using the timestamp of the run.The "train" split is always pointing to the latest… See the full description on the dataset page: https://huggingface.co/datasets/open-llm-leaderboard/NousResearch__Hermes-3-Llama-3.2-3B-details.reward-bench-Llama-3.2-1B-Instruct-yes-noGSM8K-Aug-Llama-3.2-1B-Instruct-Correct-CoT
Verified self-generated GSM8K reasoning
64 independently sampled completions are generated per prepared question.
Final answers are checked against the source answer. Among complete, correctly
formatted correct completions whose CoT passes the final-result-statement and
combined length checks, one sample is selected uniformly at random using a
reproducible per-question seed. CoT length does not rank eligible samples.
The final result belongs
in the separate final-answer line of… See the full description on the dataset page: https://huggingface.co/datasets/hanseungwook/GSM8K-Aug-Llama-3.2-1B-Instruct-Correct-CoT.
