datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
quantum-like-attention-framework-1.3b-untuned-validation
Quantum Like Attention Framework (Q.L.A.F) 1.3b untuned
This repository contains the model checkpoints, downstream evaluation scores, and pretraining convergence logs for the Quantum Like Attention Framework (Q.L.A.F) 1.3B configuration.
Key Specifications & Architecture
Model Name: Q.L.A.F 1.3b untuned (Quantum Like Attention Framework - Hybrid Architecture)
Parameters: 1.3B parameters total configuration (327M active parameter student subset)
Layer Count: 12… See the full description on the dataset page: https://huggingface.co/datasets/IgnisCogitationis/quantum-like-attention-framework-1.3b-untuned-validation.gspc-regulatory-framework
GSPC — regulatory framework facts (RegimeFacts)
SWIFT census (live): https://councilof.ai/api/swift
XRPL reader (live): https://councilof.ai/api/xrpl
MEASURED financial/domain axis (declaration presence on retrieved URLs over live XRPL reader-16, n=16). Not a model leaderboard. No accuracy, no fleet, no leader.
Live status is the regulatory-framework row on GET https://councilof.ai/api/gspc. Not a certificate.
Tokenisation evidence question (24 September 2026): What can an… See the full description on the dataset page: https://huggingface.co/datasets/csoai/gspc-regulatory-framework.Real-UI-Clickboxes
RUC: Real UI Clickboxes
Click carefully, even when the page is trying to trick you! 👀
Official Hugging Face release for RUC: Real UI Clickboxes, the dataset accompanying our ACL 2026 paper Don't Click That: Teaching Web Agents to Resist Deceptive Interfaces on deceptive UI understanding for web agents.
ACL Anthology: https://aclanthology.org/2026.acl-long.310/
PDF: https://aclanthology.org/2026.acl-long.310.pdf
DOI: https://doi.org/10.18653/v1/2026.acl-long.310… See the full description on the dataset page: https://huggingface.co/datasets/DUDE-Framework/Real-UI-Clickboxes.GEO-Framework
NobleJackal GEO Framework
A practical framework for making organisations clear, verifiable and citable in AI search
GEO means Generative Engine Optimization: the work of helping generative search and answer systems find, understand and support claims about organisations, people and content. This six-language book provides a seven-layer method for auditing entity clarity, evidence quality, machine-readable structure, question coverage, multilingual parity and… See the full description on the dataset page: https://huggingface.co/datasets/NobleJackal/GEO-Framework.Software-Architectural-FrameworksSoftware-Architectural-Frameworks
I am releasing a small dataset covering topics related to Frameworks under Software-Architecture.
I have included following topics:
TOGAF
Zachman Framework
IEEE 1471
Matrix-based approach to architecture development
Significance of IEEE 1471 (ISO/IEC 42010)
Benefits of employing architectural frameworks
and Many More!
This dataset can be useful in LLM development. Also those who are working on developing Software development related LLMs then this dataset can… See the full description on the dataset page: https://huggingface.co/datasets/ajibawa-2023/Software-Architectural-Frameworks.r9-research-framework
R9 Research Framework — Qwen3.5-9B Distillation
⚠️ CRITICAL: READ FIRST — Ollama Inference Flag Required
If you serve any Qwen3.5-derived model from this lineage via Ollama,
you MUST pass "think": false in the /api/chat request body.
curl -X POST http://localhost:11434/api/chat \
-d '{"model": "qwen3.5-9b-r10:q4km", "think": false, "messages": [...], "stream": false}'
Without this flag the model will appear to "loop" and produce empty answers
on 25-46% of requests.… See the full description on the dataset page: https://huggingface.co/datasets/cudabenchmarktest/r9-research-framework.Mitre_Attacks_Framework_Dataset
MITRE ATT&CK Enterprise Dataset
Overview
This dataset provides a comprehensive collection of MITRE ATT&CK Enterprise techniques (v14.1) in JSONL format, designed for cybersecurity professionals, red teams, and threat hunters.
Each entry maps to a specific ATT&CK technique, including its ID, name, description, real-world example, and source.
The dataset is structured for seamless integration into security tools such as SIEMs, threat intelligence platforms, or custom red… See the full description on the dataset page: https://huggingface.co/datasets/darkknight25/Mitre_Attacks_Framework_Dataset.Frameworker_User_Studyrepro-score-a-unified-framework-for-overshoot-refund-in-online-fdr-control-traces
Agent traces
Agent sessions published from a Trackio Logbook.
grounded-behavior-framework-v1_5
Grounded Behavior Framework N1 v1.5
Dataset sintético em português europeu para treino e avaliação de respostas
fundamentadas num contexto fornecido. Cada exemplo contém um contexto, uma
pergunta e uma resposta curta que aparece literalmente no contexto.
Como carregar
from datasets import load_dataset
dataset = load_dataset("empgces/grounded-behavior-framework-v1_5")
print(dataset)
print(dataset["train"][0])
Splits
Split
Exemplos
Utilização… See the full description on the dataset page: https://huggingface.co/datasets/empgces/grounded-behavior-framework-v1_5.business-frameworks
Business Frameworks
Operating judgement for running a business, distilled for agents. Query it like a consultant, pay per answer.
Scope: built for businesses from launch to about $50M in revenue; larger businesses are product two.
One document, written by an operator, for an agent that is running a business — and for an agent advising the human who does. It is structured so an agent reads the free top layer here and pays only for the node it needs; every leg ends in decision… See the full description on the dataset page: https://huggingface.co/datasets/Matryoshka-Paradigms/business-frameworks.Prettybird-Framework
🚀 The Future Standard / Geleceğin Standartı
[English]
Beyond Raw Data: The Behavioral Revolution
The AI industry has been obsessed with the volume of data. At Prometech A.Ş., we are shifting the focus to the process of thought. BCE-Prettybird-Micro-Standart is not just a collection of Q&As; it is a blueprint for behavioral reasoning. By integrating Path Mapping and Behavioral DNA into the training loop, we are setting the new industry standard: Small models with elite… See the full description on the dataset page: https://huggingface.co/datasets/pthinc/Prettybird-Framework.redteam-framework-benchmark
ORQ Red-Teaming Framework Benchmark
Overview
This dataset contains the full results of a comparative red-teaming benchmark evaluating three
open-source red-teaming frameworks — EvaluatorQ, DeepTeam, and PromptFoo — against
three victim LLMs across three target configurations and five OWASP LLM Top 10 (2025) vulnerability
categories.
Each row is one attack attempt: the attack prompt sent to the victim model, the model's response,
and the verdict from a 3-model… See the full description on the dataset page: https://huggingface.co/datasets/orq/redteam-framework-benchmark.repro-stellar-testing-framework-traces
Agent traces
Agent sessions published from a Trackio Logbook.
repro-learning-the-best-under-constraints-a-duality-based-framework-traces
Agent traces
Agent sessions published from a Trackio Logbook.
alignment-constraint-framework
Alignment Constraint Framework v1.0.0 — machine-ingestion mirror
Distribution mirror of The Alignment Constraint Framework v1.0.0.Canonical source: https://alignmentconstraint.org/Permanent framework record: https://doi.org/10.5281/zenodo.21895924GitHub release: https://github.com/bethediamond/alignment-constraint/releases/tag/v1.0.0Proof status: Stage 4 — candidate proof architecture under named premises, without independent specialist verification and without theorem… See the full description on the dataset page: https://huggingface.co/datasets/diamondlight/alignment-constraint-framework.dd-framework
📋 Due Diligence Framework
Core methodology, checklists, and templates for AI-powered due diligence analysis
This repository contains the foundational framework components for systematic due diligence analysis, including comprehensive checklists, structured question templates, and strategic analysis methodologies.
🎯 What's Included
📑 Due Diligence Checklists (2 files)
Comprehensive checklists covering all aspects of M&A due diligence:
original.md: 244 lines… See the full description on the dataset page: https://huggingface.co/datasets/jmzlx/dd-framework.Mitre_Attacks_Framework_Dataset
MITRE ATT&CK Enterprise Dataset
Overview
This dataset provides a comprehensive collection of MITRE ATT&CK Enterprise techniques (v14.1) in JSONL format, designed for cybersecurity professionals, red teams, and threat hunters.
Each entry maps to a specific ATT&CK technique, including its ID, name, description, real-world example, and source.
The dataset is structured for seamless integration into security tools such as SIEMs, threat intelligence platforms, or custom red… See the full description on the dataset page: https://huggingface.co/datasets/JR87/Mitre_Attacks_Framework_Dataset.adaption-bdl-framework
This dataset is a remastered version prepared using Adaption's Adaptive Data platform.
adaption-bdl_framework
This dataset outlines the Biosecurity Data Level (BDL) Framework for managing biological and pathogen-related data with varying security controls. It defines access tiers ranging from public to highly sensitive to prevent the misuse of dual-use research information. The content includes protocols for data tagging, access governance, and secure handling during AI… See the full description on the dataset page: https://huggingface.co/datasets/joduor/adaption-bdl-framework.ai-agents-frameworks
AI Agents Frameworks Agent Meta and Traffic Dataset in AI Agent Marketplace | AI Agent Directory | AI Agent Index from DeepNLP
This dataset is collected from AI Agent Marketplace Index and Directory at http://www.deepnlp.org, which contains AI Agents's meta information such as agent's name, website, description, as well as the monthly updated Web performance metrics, including Google,Bing average search ranking positions, Github Stars, Arxiv References, etc.
The dataset is helpful… See the full description on the dataset page: https://huggingface.co/datasets/DeepNLP/ai-agents-frameworks.sft-mobile-query-framework
SFT Mobile Query Framework Dataset
Dataset Description
This dataset contains 3607 training pairs for supervised fine-tuning (SFT) of language models to parse natural language queries about mobile phones into structured JSON execution plans.
Dataset Summary
Total Examples: 3607
Format: JSONL (question-answer pairs)
Task: Query Parsing & Structured Output Generation
Domain: Mobile Phone Specifications
Language: English
Purpose
This dataset is… See the full description on the dataset page: https://huggingface.co/datasets/sujitpandey/sft-mobile-query-framework.amber-framework-knowledge-pack
Amber Framework Knowledge Pack (demo)
The demo knowledge pack dataset behind
AgentC-Consulting/knowledge-packs:
teach a small local model the Amber web framework (Crystal),
and measure whether it learned anything with a before/after eval harness.
A knowledge pack compiles a body of expertise into curated sources, schema-validated
generated training JSONL, a contamination-guarded held-out eval set, and a JSON manifest.
This repo ships the exact training data and the two 50-item… See the full description on the dataset page: https://huggingface.co/datasets/crimson-knight/amber-framework-knowledge-pack.job_frameworks_export_2025-11-10
