CoolFace
9 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01EXDai /attention-mechanism The Attention Computation — Inside the Score Companion dataset for Episode 13 of EXD: the attention mechanism. One score matrix, one softmax, one weighted blend. This notebook loads Qwen3.6-35B-A3B, runs it up to a real full-attention layer, and reconstructs the entire attention computation by hand — QK Norm, RoPE, QKᵀ, causal mask, softmax, weighted sum, gate, and output projection — verified bit-for-bit against the model's own output. 📖 Article:… See the full description on the dataset page: https://huggingface.co/datasets/EXDai/attention-mechanism.imagefeature-extractionn<1K0 likes250 downloads1mo agoHugging Face02turhancan97 /SpaRRTa-Attention SpaRRTa-Attention: Attention-Analysis Split of the SpaRRTa Benchmark SpaRRTa-Attention is the interpretability asset for the synthetic SpaRRTa benchmark. Each scene ships with per-object segmentation masks so that a frozen Visual Foundation Model's self-attention can be measured between the objects in the scene (Human / Tree / Truck), the CLS token, the background, and register tokens. 📄 Paper: arXiv:2601.11729 💻 Code: github.com/gmum/SpaRRTa (see sparrta/analysis/) 🧩 Main (synthetic)… See the full description on the dataset page: https://huggingface.co/datasets/turhancan97/SpaRRTa-Attention.imageimage-feature-extraction1K<n<10K1 likes197 downloads3mo agoHugging Face03danny2507 /attention-uq-800q-colab Attention/UQ 800-question Colab bundle A deterministic 200-question subset for each of MultiModalQA, WebQA, HotpotQA, and TAT-QA. See manifest.json for exact upstream sources, hashes, counts, and the explicitly constructed WebQA distractor setting. imagequestion-answeringn<1K0 likes124 downloads21d agoHugging Face04omidmsl /BDD-X_attentionimage10K<n<100K0 likes53 downloads1y agoHugging Face05iamseungpil /boltzmann-attention-steering-artifactsdocumentn<1K0 likes28 downloads5mo agoHugging Face06igorgenuino /face-attention-focus-on-eyeimagen<1K0 likes4 downloads1y agoHugging Face07thotik /arithmetic-attention-h100-results Arithmetic Attention H100 Results Run: 20260401_000525 Paper: Canavesi (2026), "Semiprime Bottlenecks in Arithmetic Graphs" Configuration GPU: H100 80GB d_model: 128, heads: 4, layers: 4 Sequence lengths: up to 2048 Seeds per experiment: 2 Experiments Mask structure and sparsity scaling BFS routing efficiency (diameter analysis) Distance-dependent accuracy (scaled) Wall-clock speedup with Triton sparse kernel Gradient-based position importance… See the full description on the dataset page: https://huggingface.co/datasets/thotik/arithmetic-attention-h100-results.imagen<1K0 likes3 downloads6mo agoHugging Face08Zongrong /Attention_Diffimage1K<n<10K0 likes1 downloads2y agoHugging Face09introvoyz041 /keras-attentionimagen<1K0 likes1 downloads1y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.