datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
Long-Horizon-Terminal-Bench
Long-Horizon Terminal-Bench (LHTB)
LHTB is a 46-task benchmark for measuring how well LLM agents sustain useful
work in a containerized terminal over hundreds of steps. Unlike short-horizon
coding benchmarks where an agent writes one artifact and stops, LHTB drops the agent
into a stateful environment and grades it with hidden, rebuild-from-artifact
verifiers — self-reported progress does not count.
📝 Blog: https://zli12321.github.io/LHTB/
🏆 Leaderboard:… See the full description on the dataset page: https://huggingface.co/datasets/IntelligenceLab/Long-Horizon-Terminal-Bench.nuclear-intelligence-dataset
Nuclear Intelligence Dataset
Public, auto-generated dataset of validated nuclear-energy research cycles.
Latest stats (auto-updated):
🪙 NES tokens minted: 0
⛓️ Blockchain length: 1 blocks
🕸️ Knowledge entities: 2
Source
GitHub: https://github.com/QalamHipHop/nuclear-intelligence
HF Space: https://huggingface.co/spaces/Qalam/Nuclear-Intelligence
License
MIT
code-propriete-intellectuelle
Code de la propriété intellectuelle, non-instruct (2025-03-10)
The objective of this project is to provide researchers, professionals and law students with simplified, up-to-date access to all French legal texts, enriched with a wealth of data to facilitate their integration into Community and European projects.
Normally, the data is refreshed daily on all legal codes, and aims to simplify the production of training sets and labeling pipelines for the development of free… See the full description on the dataset page: https://huggingface.co/datasets/louisbrulenaudet/code-propriete-intellectuelle.sample-fusion-intelligence-traces
Sample Fusion Intelligence Traces
Structured AI reasoning traces from dFusion's Fusion Intelligence system. Each record captures a complete agentic workflow: a real user query on a domain-specific topic, the full message chain including system prompts, tool calls, search results, intermediate reasoning steps, and a final synthesized answer — along with human feedback.
These are not synthetic benchmarks. They are traces from real queries submitted by real users on live financial… See the full description on the dataset page: https://huggingface.co/datasets/dFusionAILabs/sample-fusion-intelligence-traces.intellect-3-rl-math-5k
intellect-3-rl-math-5k
A difficulty-stratified sample of 5,000 unique math problems drawn from PrimeIntellect/INTELLECT-3-RL.
How it was drawn
Source pool: 12kimih/intellect-3-rl-math-decontaminated, 20,918 unique problems, the math config, decontaminated against the benchmarks listed below.
Stratum: stratum, the number of 8 attempts by Qwen3-4B-Thinking-2507 that matched the reference answer, shipped per problem by the upstream. It runs 0 (never solved) to 8… See the full description on the dataset page: https://huggingface.co/datasets/12kimih/intellect-3-rl-math-5k.intellect-3-rl-math-decontaminated
intellect-3-rl-math-decontaminated
20,918 unique math problems from PrimeIntellect/INTELLECT-3-RL with 243 removed as contaminated.
How it was prepared
Pool: the math config of INTELLECT-3-RL, 21,161 rows, normalised to the column names used here.
Rows: 21,161 loaded, 20,918 kept.
Decontaminated against: math500, aime2024, aime2025, aime2026, amc, hmmt_feb2023, hmmt_feb2024, hmmt_feb2025, hmmt_feb2026, hmmt_nov2025, olympiadbench, gsm8k (243 problems removed… See the full description on the dataset page: https://huggingface.co/datasets/12kimih/intellect-3-rl-math-decontaminated.Turkish-Lyric-Intelligence-v2
Turkish Lyric Intelligence v2
Turkish Lyric Intelligence v2 is a derived, research-oriented dataset for
building controllable Turkish lyric generation and lyric-analysis systems. It
adds section structure, orthographic syllable counts, rhyme-ending candidates,
repetition signals, and review queues to the source
Genius Turkish Dataset.
This release is intended as an intermediate data layer. Automated rhyme,
prosody, and task labels are candidate annotations, not expert literary… See the full description on the dataset page: https://huggingface.co/datasets/mustafakemal0146/Turkish-Lyric-Intelligence-v2.ai-paper-intellectual-lineage-2023
Intellectual Lineage of Impactful AI Research Papers (2023-2024)
Dataset Description
This dataset contains 20 impactful AI research papers published between 2022-2024, along with their intellectual lineage - tracing 1-2 key prior works each paper builds upon, and a ~300-word paragraph explaining the relationship between the current work and its foundations.
Purpose
Understanding how research ideas evolve and build upon prior work is crucial for:
Researchers… See the full description on the dataset page: https://huggingface.co/datasets/AmberLJC/ai-paper-intellectual-lineage-2023.SwarmFailure-Intelligence
SwarmFailure-Intelligence v1
A dataset of real AI system failures, diagnoses, and repair strategies.
SwarmFailure-Intelligence is the first structured reliability dataset purpose-built for training LLMs and agents to detect, diagnose, repair, and prevent AI system failures. Every record traces a concrete failure through its full lifecycle -- from the broken execution to root cause analysis to a validated fix.
This is not synthetic noise. Every pair was generated from agent execution… See the full description on the dataset page: https://huggingface.co/datasets/SwarmandBee/SwarmFailure-Intelligence.Emotional_Intelligence_Corpus
Emotional Intelligence
This corpus was automatically generated by the Deku Corpus Builder for use in RAG-based AI applications.
Dataset Structure
Each record contains:
text: The content text
source_url: Original source URL
source_title: Title of the source document
source_domain: Domain of the source
relevance_score: Relevance to the subject (0-1)
quality_score: Content quality score (0-1)
topics: JSON array of detected topics
character_count: Length of the text… See the full description on the dataset page: https://huggingface.co/datasets/PhillyMac/Emotional_Intelligence_Corpus.Emotional_Intelligence_Content_1
Emotional Intelligence Content 1
This corpus was automatically generated by the Deku Corpus Builder for use in RAG-based AI applications.
Dataset Structure
Each record contains:
text: The content text
source_url: Original source URL
source_title: Title of the source document
source_domain: Domain of the source
license_type: License classification (e.g. public_domain, cc_by, cc_by_sa)
attribution_required: Boolean — True for CC BY / CC BY-SA and other attribution-required… See the full description on the dataset page: https://huggingface.co/datasets/PhillyMac/Emotional_Intelligence_Content_1.Emotional_Intelligence_Content_2
Emotional Intelligence Content 2
This corpus was automatically generated by the Deku Corpus Builder for use in RAG-based AI applications.
Dataset Structure
Each record contains:
text: The content text
source_url: Original source URL
source_title: Title of the source document
source_domain: Domain of the source
license_type: License classification (e.g. public_domain, cc_by, cc_by_sa)
attribution_required: Boolean — True for CC BY / CC BY-SA and other attribution-required… See the full description on the dataset page: https://huggingface.co/datasets/PhillyMac/Emotional_Intelligence_Content_2.
