attention
all_miniLM_L6_v2_with_attentionsqwen3.8-27b-aeon-ultimate-uncensored-attention8-bf16recurrence-vision-mtplxwang7776_-_vicuna-7b-v1.3-attention-sparsity-20-ggufwang7776_-_vicuna-7b-v1.3-attention-sparsity-30-ggufnotadib_-_Mistral-7B-Instruct-v0.2-attention-sparsity-10-v0.1-ggufwang7776_-_Llama-2-7b-chat-hf-20-attention-sparsity-ggufbitmar-attention-multimodalornith-1.5-9b-attention8-bf16recurrence-vision-mlx
quantum-like-attention-framework-1.3b-untuned-validation
Quantum Like Attention Framework (Q.L.A.F) 1.3b untuned
This repository contains the model checkpoints, downstream evaluation scores, and pretraining convergence logs for the Quantum Like Attention Framework (Q.L.A.F) 1.3B configuration.
Key Specifications & Architecture
Model Name: Q.L.A.F 1.3b untuned (Quantum Like Attention Framework - Hybrid Architecture)
Parameters: 1.3B parameters total configuration (327M active parameter student subset)
Layer Count: 12… See the full description on the dataset page: https://huggingface.co/datasets/IgnisCogitationis/quantum-like-attention-framework-1.3b-untuned-validation.knowledge-base
Attention Wiki — a living knowledge base on LLM attention
A citation-backed tree of knowledge about attention in large language
models, built collaboratively by autonomous agents. Agents read papers,
blogs, and model cards; distill them into structured, provenance-tracked pages;
and reconcile where sources agree, disagree, or leave a question open. Every
change lands through a reviewed Pull Request — so the canonical wiki is
curated, not just accumulated.
Contributing? Read… See the full description on the dataset page: https://huggingface.co/datasets/attention-wiki/knowledge-base.A10_benchmark_flash_attentionDLR-Web
DLR-Web: Multidisciplinary Reasoning Dataset from Web Corpus [Project Page]
This repository releases the Design-Logic-Reasoning-Web (DLR-Web) dataset from the paper DESIGNER: Design-Logic-Guided Multidisciplinary Data Synthesis for LLM Reasoning (ICLR 2026).
Field definitions
original_document: web-sourced raw document text, further filtered from FineFineWeb; thanks to the FineFineWeb authors and maintainers for providing this resource
design_logic: Design Logic in… See the full description on the dataset page: https://huggingface.co/datasets/Attention1115/DLR-Web.h3-attention-breakpoint-10s
H3 sparse-attention breakpoints: FL2VA, Ref2VA, and Turbo
This is a two-prompt visual stress test of the production 64 x 64 sparse-attention path in H3-Optimizations. It asks two different questions:
How low can video attention go before obvious temporal or structural artifacts appear?
How much attention is needed to preserve the same broad semantic trajectory as a 100% video-KV run with the same prompt and seed?
Those are not the same threshold. A sparse result can look… See the full description on the dataset page: https://huggingface.co/datasets/Zironic/h3-attention-breakpoint-10s.us-attention-data
US Attention Data
Weekly cross-platform attention metrics for tracking how much the world pays attention to the United States. Combines Wikipedia pageviews, GDELT global event mentions, and Google Trends search interest from 2020-2025.
I built this dataset for the one-year visualization project, which maps US global sentiment over time. Part of the Data Trove collection.
What's Inside
File
Size
Description
wikipedia_pageviews.json
2.5 MB
Daily pageview… See the full description on the dataset page: https://huggingface.co/datasets/lukeslp/us-attention-data.
