CoolFace
16 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01nebius /Llama-3.1-8B-Instruct-Infinity-Instruct-0625 Llama-3.1-8B-Instruct-Infinity-Instruct-0625 Dataset Description This dataset is part of the LK-Speculators collection for speculative decoding research. It contains 660K prompt-response pairs designed for training draft models that are used alongside Llama-3.1-8B-Instruct as the target model. The dataset was created by generating responses to the prompts from Infinity-Instruct-0625 with meta-llama/Llama-3.1-8B-Instruct at temperature=1. For more details on the training… See the full description on the dataset page: https://huggingface.co/datasets/nebius/Llama-3.1-8B-Instruct-Infinity-Instruct-0625.texttext-generation100K<n<1M1 likes85 downloads7mo agoHugging Face02Transluce /input_ablation_llama_3.1_8b_instruct_mmlu_hint Training Language Models to Explain Their Own Computations - Input Ablations This dataset is part of the research presented in the paper Training Language Models to Explain Their Own Computations. It contains data for the Input Ablations task, where explainer models are trained to predict how removing input hints affects the target model's (Llama-3.1-8B-Instruct) predictions on MMLU questions with hints. This task evaluates whether models can understand the causal relationships… See the full description on the dataset page: https://huggingface.co/datasets/Transluce/input_ablation_llama_3.1_8b_instruct_mmlu_hint.texttext-generation10K<n<100K0 likes42 downloads9mo agoHugging Face03suchirsalhan /kidalign-llama-3.1-8B-Instruct KidAlign: Llama-3.1-8B-Instruct Results This dataset contains synthetic survey responses generated by Llama-3.1-8B-Instruct under various demographic and personality conditions. Dataset Structure Each entry represents a full survey session. To filter by individual questions within the Hugging Face UI, use the Search bar or the Nested Field Explorer in the Dataset Viewer. Metadata Columns (Filters) Field Description condition.age_group The target age… See the full description on the dataset page: https://huggingface.co/datasets/suchirsalhan/kidalign-llama-3.1-8B-Instruct.text-generation0 likes32 downloads5mo agoHugging Face04ZachW /llama-3.1-8b-instruct_aime-all meta-llama/Llama-3.1-8B-Instruct — aime-all Model outputs from the micro-creativity inference suite. Model: meta-llama/Llama-3.1-8B-Instruct Dataset: aime-all (933 items) Part of collection: ZachW/llm-creativity-benchmarks Generation config temperature: 0.0 max_tokens: 32768 seed: 42 backend: vllm Columns Column Description task_id Unique task identifier input The exact prompt sent to the model (after meta-prompt application)… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_aime-all.tabulartext-generationn<1K0 likes22 downloads5mo agoHugging Face05ilyasoulk /ai-vs-human-meta-llama-Llama-3.1-8B-Instruct-CNN AI vs Human dataset on the CNN DailyNews Dataset Description This dataset showcases pairs of truncated text and their respective completions, crafted either by humans or an AI language model. Each article was randomly truncated between 25% and 50% of its length. The language model was then tasked with generating a completion that mirrored the characters count of the original human-written continuation. Data Fields 'human': The original human-authored… See the full description on the dataset page: https://huggingface.co/datasets/ilyasoulk/ai-vs-human-meta-llama-Llama-3.1-8B-Instruct-CNN.texttext-classification1K<n<10K1 likes19 downloads2y agoHugging Face06ZachW /llama-3.1-8b-instruct_ifeval meta-llama/Llama-3.1-8B-Instruct — ifeval Model outputs from the micro-creativity inference suite. Model: meta-llama/Llama-3.1-8B-Instruct Dataset: ifeval (541 items) Part of collection: ZachW/llm-creativity-benchmarks Generation config temperature: 0.0 max_tokens: 16384 seed: 42 backend: vllm Columns Column Description task_id Unique task identifier input The exact prompt sent to the model (after meta-prompt application)… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_ifeval.tabulartext-generationn<1K0 likes19 downloads5mo agoHugging Face07codezakh /EFAGen-Llama-3.1-8B-Instruct-Training-DataPaper Link The training data used for the final version of EFAGen-Llama-3.1-8B-Instruct. The data is in Alpaca format and can be used with Llama-Factory (check dataset_info.json). texttext-generation1K<n<10K1 likes16 downloads1y agoHugging Face08sapbot /llama-3.1-8b-instruct-419xTrace of Llama 3.1 8B Instruct LLM by Meta. Data count (Total: 419): English - 204 Russian - 215 Data is presented in ChatML format and each conversation split by newline. P.S. I do not recommend anyone to use data from this model, because it's.... way too stupid and random at some time, results are not distill-worthy. Go check Gemma 3 12B dataset. This model was NOT free, and I had to use OpenRouter for it. Crypto donations for future projects like this are available on my personal page texttext-generationn<1K0 likes16 downloads5mo agoHugging Face09ZachW /llama-3.1-8b-instruct_arena-hard-creative-writing meta-llama/Llama-3.1-8B-Instruct — arena-hard-creative-writing Model outputs from the micro-creativity inference suite. Model: meta-llama/Llama-3.1-8B-Instruct Dataset: arena-hard-creative-writing (250 items) Part of collection: ZachW/llm-creativity-benchmarks Generation config temperature: 0.0 max_tokens: 16384 seed: 42 backend: vllm Columns Column Description task_id Unique task identifier input The exact prompt sent to the… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_arena-hard-creative-writing.tabulartext-generationn<1K0 likes14 downloads5mo agoHugging Face10ZachW /llama-3.1-8b-instruct_bookmia-label0-5pct-raw meta-llama/Llama-3.1-8B-Instruct — bookmia-label0-5pct-raw Model outputs from the micro-creativity inference suite. Model: meta-llama/Llama-3.1-8B-Instruct Dataset: bookmia-label0-5pct-raw (247 items) Part of collection: ZachW/llm-creativity-benchmarks Generation config temperature: 0.0 max_tokens: 16384 seed: 42 backend: vllm Columns Column Description task_id Unique task identifier input The exact prompt sent to the model… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_bookmia-label0-5pct-raw.tabulartext-generationn<1K0 likes12 downloads5mo agoHugging Face11ZachW /llama-3.1-8b-instruct_storygen-prompts-200 meta-llama/Llama-3.1-8B-Instruct — storygen-prompts-200 Model outputs from the micro-creativity inference suite. Model: meta-llama/Llama-3.1-8B-Instruct Dataset: storygen-prompts-200 (200 items) Part of collection: ZachW/llm-creativity-benchmarks Generation config temperature: 0.0 max_tokens: 16384 seed: 42 backend: vllm Columns Column Description task_id Unique task identifier input The exact prompt sent to the model (after… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_storygen-prompts-200.tabulartext-generationn<1K0 likes12 downloads5mo agoHugging Face12ZachW /llama-3.1-8b-instruct_writingbench-en100 meta-llama/Llama-3.1-8B-Instruct — writingbench-en100 Model outputs from the micro-creativity inference suite. Model: meta-llama/Llama-3.1-8B-Instruct Dataset: writingbench-en100 (100 items) Part of collection: ZachW/llm-creativity-benchmarks Generation config temperature: 0.0 max_tokens: 8192 seed: 42 backend: vllm Columns Column Description task_id Unique task identifier input The exact prompt sent to the model (after… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_writingbench-en100.tabulartext-generationn<1K0 likes12 downloads5mo agoHugging Face13ilyasoulk /ai-vs-human-meta-llama-Llama-3.1-8B-Instruct AI vs Human dataset on the OpenWebTxt Dataset Description This dataset showcases pairs of truncated text and their respective completions, crafted either by humans or an AI language model. Each article was randomly truncated between 25% and 50% of its length. The language model was then tasked with generating a completion that mirrored the characters count of the original human-written continuation. Data Fields 'human': The original human-authored… See the full description on the dataset page: https://huggingface.co/datasets/ilyasoulk/ai-vs-human-meta-llama-Llama-3.1-8B-Instruct.texttext-classification1K<n<10K1 likes11 downloads2y agoHugging Face14ZachW /llama-3.1-8b-instruct_creativemath-with-answers meta-llama/Llama-3.1-8B-Instruct — creativemath-with-answers Model outputs from the micro-creativity inference suite. Model: meta-llama/Llama-3.1-8B-Instruct Dataset: creativemath-with-answers (188 items) Part of collection: ZachW/llm-creativity-benchmarks Generation config temperature: 0.0 max_tokens: 32768 seed: 42 backend: vllm Columns Column Description task_id Unique task identifier input The exact prompt sent to the… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_creativemath-with-answers.tabulartext-generationn<1K0 likes11 downloads5mo agoHugging Face15ZachW /llama-3.1-8b-instruct_tinystories-val1pct-raw meta-llama/Llama-3.1-8B-Instruct — tinystories-val1pct-raw Model outputs from the micro-creativity inference suite. Model: meta-llama/Llama-3.1-8B-Instruct Dataset: tinystories-val1pct-raw (220 items) Part of collection: ZachW/llm-creativity-benchmarks Generation config temperature: 0.0 max_tokens: 16384 seed: 42 backend: vllm Columns Column Description task_id Unique task identifier input The exact prompt sent to the model… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_tinystories-val1pct-raw.tabulartext-generationn<1K0 likes11 downloads5mo agoHugging Face16ZachW /llama-3.1-8b-instruct_alpaca-text-generation-384 meta-llama/Llama-3.1-8B-Instruct — alpaca-text-generation-384 Model outputs from the micro-creativity inference suite. Model: meta-llama/Llama-3.1-8B-Instruct Dataset: alpaca-text-generation-384 (384 items) Part of collection: ZachW/llm-creativity-benchmarks Generation config temperature: 0.0 max_tokens: 16384 seed: 42 backend: vllm Columns Column Description task_id Unique task identifier input The exact prompt sent to the… See the full description on the dataset page: https://huggingface.co/datasets/ZachW/llama-3.1-8b-instruct_alpaca-text-generation-384.tabulartext-generationn<1K0 likes9 downloads5mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.