CoolFace
22 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01astro-legacy-archive /gaia-dr3-oa-neuron-xp-spectra Gaia DR3 OA neuron XP spectra This table contains the prototype BP/RP spectrum attached to each neuron in the 30 × 30 self-organising map produced by Gaia's Apsis Outlier Analysis module. Measured from the served table, every (neuron_id, xp_spectrum_prototype_wavelength) pair is unique across its 78,300 rows. Each of the 900 neurons has the same ordered grid of 87 wavelengths from 379.0935 to 1034.2460 nm. Neuron-level statistics and other attributes belong to Gaia's separate… See the full description on the dataset page: https://huggingface.co/datasets/astro-legacy-archive/gaia-dr3-oa-neuron-xp-spectra.tabular10K<n<100K0 likes111 downloads2d agoHugging Face02hbe /neuronpedia-sae-concepts Neuronpedia SAE Concepts Complete extraction of all individual concepts from every Sparse Autoencoder (SAE) released on Neuronpedia, plus all public features from Anthropic's Towards Monosemanticity (2023) and Scaling Monosemanticity (2024) papers. Quick Start from datasets import load_dataset # Full Neuronpedia dataset (77M rows, streaming recommended) ds = load_dataset("hbe/neuronpedia-sae-concepts", split="train", streaming=True) # Unique concepts with essential… See the full description on the dataset page: https://huggingface.co/datasets/hbe/neuronpedia-sae-concepts.tabularfeature-extraction100M<n<1B0 likes99 downloads4mo agoHugging Face03Confirm-Labs /pythia-12b-neuron-dataset-examples pythia-12b-neuron-dataset-examples This dataset contains the top 64 highest activating dataset examples for each MLP neuron in Pythia-12b. The dataset examples are all 16 tokens long. See https://confirmlabs.org/posts/dreaming.html for details. Columns: layer: the layer of the neuron neuron: the index of the neuron rank: the rank of the example activation: the activation of the neuron on the example position: the token position for which the neuron is maximally activated. text: the… See the full description on the dataset page: https://huggingface.co/datasets/Confirm-Labs/pythia-12b-neuron-dataset-examples.tabular10M<n<100M2 likes83 downloads3y agoHugging Face04Neuronovo /neuronovo-utc-data-glue-mnlitabular1M<n<10M0 likes65 downloads3y agoHugging Face05neuronetties /coffee_sales_datatabular100K<n<1M1 likes53 downloads2y agoHugging Face06generative-latent-prior /llama8b-layer15-meta-neurons Llama8B Meta-Neurons This repository contains meta-neuron data accompanying the paper Learning a Generative Meta-Model of LLM Activations. Project page: https://generative-latent-prior.github.io Code: https://github.com/g-luo/generative_latent_prior Quick Start With this data, you can browse the 98304 meta-neurons of the Llama-3.1-8B GLP (glp-llama8b-d6, Layer 15). Meta-neurons are the post-SwiGLU activations of the GLP's MLP blocks. For each meta-neuron… See the full description on the dataset page: https://huggingface.co/datasets/generative-latent-prior/llama8b-layer15-meta-neurons.tabular10K<n<100K0 likes53 downloads12d agoHugging Face07Neuronovo /neuronovo-utc-data-goemotionstabular100K<n<1M1 likes52 downloads3y agoHugging Face08Neuronovo /neuronovo-utc-persent-doctabular10K<n<100K0 likes39 downloads2y agoHugging Face09Neuronovo /neuronovo-utc-data-glue-colatabular1K<n<10K0 likes34 downloads3y agoHugging Face10sjgerstner /Llama-3.2-3B_neuron-activationstabular100K<n<1M0 likes28 downloads2mo agoHugging Face11sjgerstner /Llama-3.1-8B_neuron-activationstabular100K<n<1M0 likes26 downloads2mo agoHugging Face12Neuronovo /neuronovo-utc-tweeteval-sentimenttabular100K<n<1M0 likes25 downloads2y agoHugging Face13Neuronovo /neuronovo-utc-unhealthy-conversationstabular100K<n<1M0 likes24 downloads2y agoHugging Face14Neuronovo /neuronovo-utc-measuring-hate-speechtabular100K<n<1M0 likes24 downloads2y agoHugging Face15Neuronovo /argilla_preferences_personalized_filteredtabular10K<n<100K0 likes22 downloads2y agoHugging Face16Neuronovo /rlhf_synthetic_generalizedtabularn<1K0 likes21 downloads2y agoHugging Face17Neuronovo /neuronovo-utc-tweeteval-emotionstabular10K<n<100K0 likes19 downloads2y agoHugging Face18Neuronovo /neuronovo-utc-hate-speech18-sentencestabular10K<n<100K0 likes18 downloads2y agoHugging Face19almogtavor /amortized-neuron-pruning-effects Amortized Neuron-Ablation Effects (for pruning) Exact mean-ablation effects of individual MLP neurons in transformer LMs, paired with cheap forward/backward signals, for the task of amortized causal intervention-effect prediction at neuron granularity (pruning). The intervention unit is a single MLP neuron (blocks.{l}.mlp.hook_post channel). For every (prompt, neuron) we record the exact effect of replacing that neuron's last-token activation with its dataset mean, m(x… See the full description on the dataset page: https://huggingface.co/datasets/almogtavor/amortized-neuron-pruning-effects.tabular10M<n<100M0 likes15 downloads3mo agoHugging Face20sjgerstner /OLMo-7B-0424-hf_neuron-activationsThis dataset contains activation data of neurons in OLMo-7B-0424. (We define a neuron as a hidden dimension in a MLP sublayer.) To create the dataset, the model was run on 20M tokens from Dolma (in the same collection, we also release the Dolma subset, which we call dolma-small). Dataset Description Each row corresponds to a neuron, identified by the columns "layer" and "neuron". (We use zero-based indexing). The other columns are as follows: The first two elements of the name… See the full description on the dataset page: https://huggingface.co/datasets/sjgerstner/OLMo-7B-0424-hf_neuron-activations.tabulartabular-regression100K<n<1M0 likes13 downloads2mo agoHugging Face21eerwitt /qwen-h-neurons-datasettabular1K<n<10K0 likes9 downloads7mo agoHugging Face22sjgerstner /gemma-2-2b_neuron-activationstabular100K<n<1M0 likes6 downloads2mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.