datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Llama-3.1-8B-Instruct-Infinity-Instruct-0625
Llama-3.1-8B-Instruct-Infinity-Instruct-0625
Dataset Description
This dataset is part of the LK-Speculators collection for speculative decoding research. It contains 660K prompt-response pairs designed for training draft models that are used alongside Llama-3.1-8B-Instruct as the target model. The dataset was created by generating responses to the prompts from Infinity-Instruct-0625 with meta-llama/Llama-3.1-8B-Instruct at temperature=1.
For more details on the training… See the full description on the dataset page: https://huggingface.co/datasets/nebius/Llama-3.1-8B-Instruct-Infinity-Instruct-0625.input_ablation_llama_3.1_8b_instruct_mmlu_hint
Training Language Models to Explain Their Own Computations - Input Ablations
This dataset is part of the research presented in the paper Training Language Models to Explain Their Own Computations.
It contains data for the Input Ablations task, where explainer models are trained to predict how removing input hints affects the target model's (Llama-3.1-8B-Instruct) predictions on MMLU questions with hints. This task evaluates whether models can understand the causal relationships… See the full description on the dataset page: https://huggingface.co/datasets/Transluce/input_ablation_llama_3.1_8b_instruct_mmlu_hint.kidalign-llama-3.1-8B-Instruct
KidAlign: Llama-3.1-8B-Instruct Results
This dataset contains synthetic survey responses generated by Llama-3.1-8B-Instruct under various demographic and personality conditions.
Dataset Structure
Each entry represents a full survey session. To filter by individual questions within the Hugging Face UI, use the Search bar or the Nested Field Explorer in the Dataset Viewer.
Metadata Columns (Filters)
Field
Description
condition.age_group
The target age… See the full description on the dataset page: https://huggingface.co/datasets/suchirsalhan/kidalign-llama-3.1-8B-Instruct.llama-3.1-8b-instruct_aime-all
meta-llama/Llama-3.1-8B-Instruct — aime-all
Model outputs from the micro-creativity inference suite.
Model: meta-llama/Llama-3.1-8B-Instruct
Dataset: aime-all (933 items)
Part of collection: ZachW/llm-creativity-benchmarks
Generation config
temperature: 0.0
max_tokens: 32768
seed: 42
backend: vllm
Columns
Column
Description
task_id
Unique task identifier
input
The exact prompt sent to the model (after meta-prompt application)… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_aime-all.ai-vs-human-meta-llama-Llama-3.1-8B-Instruct-CNN
AI vs Human dataset on the CNN DailyNews
Dataset Description
This dataset showcases pairs of truncated text and their respective completions, crafted either by humans or an AI language model.
Each article was randomly truncated between 25% and 50% of its length.
The language model was then tasked with generating a completion that mirrored the characters count of the original human-written continuation.
Data Fields
'human': The original human-authored… See the full description on the dataset page: https://huggingface.co/datasets/ilyasoulk/ai-vs-human-meta-llama-Llama-3.1-8B-Instruct-CNN.llama-3.1-8b-instruct_ifeval
meta-llama/Llama-3.1-8B-Instruct — ifeval
Model outputs from the micro-creativity inference suite.
Model: meta-llama/Llama-3.1-8B-Instruct
Dataset: ifeval (541 items)
Part of collection: ZachW/llm-creativity-benchmarks
Generation config
temperature: 0.0
max_tokens: 16384
seed: 42
backend: vllm
Columns
Column
Description
task_id
Unique task identifier
input
The exact prompt sent to the model (after meta-prompt application)… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_ifeval.EFAGen-Llama-3.1-8B-Instruct-Training-DataPaper Link
The training data used for the final version of EFAGen-Llama-3.1-8B-Instruct.
The data is in Alpaca format and can be used with Llama-Factory (check dataset_info.json).
llama-3.1-8b-instruct-419xTrace of Llama 3.1 8B Instruct LLM by Meta.
Data count (Total: 419):
English - 204
Russian - 215
Data is presented in ChatML format and each conversation split by newline.
P.S. I do not recommend anyone to use data from this model, because it's.... way too stupid and random at some time, results are not distill-worthy. Go check Gemma 3 12B dataset.
This model was NOT free, and I had to use OpenRouter for it. Crypto donations for future projects like this are available on my personal page
llama-3.1-8b-instruct_arena-hard-creative-writing
meta-llama/Llama-3.1-8B-Instruct — arena-hard-creative-writing
Model outputs from the micro-creativity inference suite.
Model: meta-llama/Llama-3.1-8B-Instruct
Dataset: arena-hard-creative-writing (250 items)
Part of collection: ZachW/llm-creativity-benchmarks
Generation config
temperature: 0.0
max_tokens: 16384
seed: 42
backend: vllm
Columns
Column
Description
task_id
Unique task identifier
input
The exact prompt sent to the… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_arena-hard-creative-writing.llama-3.1-8b-instruct_bookmia-label0-5pct-raw
meta-llama/Llama-3.1-8B-Instruct — bookmia-label0-5pct-raw
Model outputs from the micro-creativity inference suite.
Model: meta-llama/Llama-3.1-8B-Instruct
Dataset: bookmia-label0-5pct-raw (247 items)
Part of collection: ZachW/llm-creativity-benchmarks
Generation config
temperature: 0.0
max_tokens: 16384
seed: 42
backend: vllm
Columns
Column
Description
task_id
Unique task identifier
input
The exact prompt sent to the model… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_bookmia-label0-5pct-raw.llama-3.1-8b-instruct_storygen-prompts-200
meta-llama/Llama-3.1-8B-Instruct — storygen-prompts-200
Model outputs from the micro-creativity inference suite.
Model: meta-llama/Llama-3.1-8B-Instruct
Dataset: storygen-prompts-200 (200 items)
Part of collection: ZachW/llm-creativity-benchmarks
Generation config
temperature: 0.0
max_tokens: 16384
seed: 42
backend: vllm
Columns
Column
Description
task_id
Unique task identifier
input
The exact prompt sent to the model (after… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_storygen-prompts-200.llama-3.1-8b-instruct_writingbench-en100
meta-llama/Llama-3.1-8B-Instruct — writingbench-en100
Model outputs from the micro-creativity inference suite.
Model: meta-llama/Llama-3.1-8B-Instruct
Dataset: writingbench-en100 (100 items)
Part of collection: ZachW/llm-creativity-benchmarks
Generation config
temperature: 0.0
max_tokens: 8192
seed: 42
backend: vllm
Columns
Column
Description
task_id
Unique task identifier
input
The exact prompt sent to the model (after… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_writingbench-en100.ai-vs-human-meta-llama-Llama-3.1-8B-Instruct
AI vs Human dataset on the OpenWebTxt
Dataset Description
This dataset showcases pairs of truncated text and their respective completions, crafted either by humans or an AI language model.
Each article was randomly truncated between 25% and 50% of its length.
The language model was then tasked with generating a completion that mirrored the characters count of the original human-written continuation.
Data Fields
'human': The original human-authored… See the full description on the dataset page: https://huggingface.co/datasets/ilyasoulk/ai-vs-human-meta-llama-Llama-3.1-8B-Instruct.llama-3.1-8b-instruct_creativemath-with-answers
meta-llama/Llama-3.1-8B-Instruct — creativemath-with-answers
Model outputs from the micro-creativity inference suite.
Model: meta-llama/Llama-3.1-8B-Instruct
Dataset: creativemath-with-answers (188 items)
Part of collection: ZachW/llm-creativity-benchmarks
Generation config
temperature: 0.0
max_tokens: 32768
seed: 42
backend: vllm
Columns
Column
Description
task_id
Unique task identifier
input
The exact prompt sent to the… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_creativemath-with-answers.llama-3.1-8b-instruct_tinystories-val1pct-raw
meta-llama/Llama-3.1-8B-Instruct — tinystories-val1pct-raw
Model outputs from the micro-creativity inference suite.
Model: meta-llama/Llama-3.1-8B-Instruct
Dataset: tinystories-val1pct-raw (220 items)
Part of collection: ZachW/llm-creativity-benchmarks
Generation config
temperature: 0.0
max_tokens: 16384
seed: 42
backend: vllm
Columns
Column
Description
task_id
Unique task identifier
input
The exact prompt sent to the model… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_tinystories-val1pct-raw.llama-3.1-8b-instruct_alpaca-text-generation-384
meta-llama/Llama-3.1-8B-Instruct — alpaca-text-generation-384
Model outputs from the micro-creativity inference suite.
Model: meta-llama/Llama-3.1-8B-Instruct
Dataset: alpaca-text-generation-384 (384 items)
Part of collection: ZachW/llm-creativity-benchmarks
Generation config
temperature: 0.0
max_tokens: 16384
seed: 42
backend: vllm
Columns
Column
Description
task_id
Unique task identifier
input
The exact prompt sent to the… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_alpaca-text-generation-384.
