datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
dojo_market_dynamicsSTUZero-Atari-Dynamics
STUZero Atari Dynamics Dataset
Offline dynamics training datasets collected from trained EfficientZero V2 (EZv2) benchmark models on Atari games. Each game's data is stored in a subfolder named {game}_{steps} indicating the game and the number of training steps of the source checkpoint. While all models were trained for 120K steps, best results in some games were attained at earlier checkpoints. The model with best eval scores was used to curate data for each game.… See the full description on the dataset page: https://huggingface.co/datasets/Shivamkak/STUZero-Atari-Dynamics.humanoid-robots-training-dataset
Dynamic Intelligence — Humanoid Robot Training Dataset
A first-person (egocentric) video dataset of human hand manipulation, designed for training humanoid robot policies via imitation learning. Each episode captures a person performing an everyday household task — folding clothes, moving dishes, opening doors — filmed from a head-mounted iPhone using its built-in LiDAR and depth sensors.
The dataset pairs each video with frame-level 3D hand tracking and camera pose data, giving… See the full description on the dataset page: https://huggingface.co/datasets/DynamicIntelligence/humanoid-robots-training-dataset.EverMemBench-Dynamic
EverMemBench-Dynamic
A benchmark dataset for evaluating long-term memory capabilities in conversational AI systems. It is part of EverMemBench, the first benchmark designed for long-horizon collaborative memory, introduced in the paper Evaluating Long-Horizon Memory for Multi-Party Collaborative Dialogues — accepted at KDD 2026 (Oral).
Configurations
This dataset has three configurations (subsets):
dialogues
Multi-turn group dialogues spanning ~250… See the full description on the dataset page: https://huggingface.co/datasets/EverMind-AI/EverMemBench-Dynamic.live-facts-snapshot
Live Facts Snapshot
A daily snapshot of verifiable, post-training-cutoff world-state facts — the kind of
ground truth language models cannot know from training data — exported through
Dynamic Feed, a live, verifiable data API whose every response
is Ed25519-signed. One file per day (data/YYYY-MM-DD.jsonl), one fact per line, and
every row carries its own source, source_url and measured_at.
Facts covered per day:
tool
facts
upstream source
licence
software_version… See the full description on the dataset page: https://huggingface.co/datasets/dynamicfeed/live-facts-snapshot.DynamicMCPBench
DynamicMCPBench
A trace-grounded, effect-scored benchmark for LLM agents on live MCP servers.
Tasks are generated forward: an explorer agent drives real MCP tools until a goal
is reached, the recorded trace is distilled into a TaskSpec, and candidates are graded
on whether they reproduce the effects the trace produced — checkpoints, equivalence
sets, minefields, a partial order — never on matching an answer string or a fixed tool
list. Candidates are evaluated under… See the full description on the dataset page: https://huggingface.co/datasets/TokenWasteGroup/DynamicMCPBench.dynamic_earthnetDynamic EarthNet dataset redistributed from https://mediatum.ub.tum.de/1650201 and https://cvg.cit.tum.de/webshare/u/toker/dynnet_training_splits/ under a common tarball for simpler download speeds.
Individual zip files were replaced with tarballs instead.
In the mediatum server version the following directories have the wrong name compared to the given split txt files:
/labels/5111_4560_13_38S
/labels/6204_3495_13_46N
/labels/7026_3201_13_52N
/labels/7367_5050_13_54S
/labels/2459_4406_13_19S… See the full description on the dataset page: https://huggingface.co/datasets/torchgeo/dynamic_earthnet.Semantic-Flow-Dynamics-SFD
Semantic Flow Dynamics (SFD) — A Formally Specified Social-Science Theory Corpus
TL;DR: 614 Chinese-language formalized social-science concepts across 25
papers, UUID-linked with typed derivation relations (derives_from,
leads_to, falsified_by, …) — usable for knowledge-graph construction,
RAG over structured theory, or as a Chinese formal-reasoning corpus.
Author: 黃正宇 Cheng Yu HuangContact: mthree.tw@gmail.com
What This Dataset Is
This corpus is an ongoing… See the full description on the dataset page: https://huggingface.co/datasets/mthreetw/Semantic-Flow-Dynamics-SFD.groundtruth-dynamic-benchmarking
Groundtruth Dynamic Benchmarking — Geology
Question sets and grading rubrics for evaluating LLMs on real-world geological
reasoning. Every question is authored from a real source corpus, and every
claim in the grading key carries an evidence locator back to that corpus —
nothing is synthetic. Licensing/redistribution status varies by corpus — see
License.
This dataset holds the questions, grading rubrics, and source corpora.
Running an evaluation (generating answers from a model… See the full description on the dataset page: https://huggingface.co/datasets/EigenformAI/groundtruth-dynamic-benchmarking.safe-alignment-dynamic
safe-alignment-dynamic
Training prompts for score-conditioned SFT / RL and separate reward-model pair sets; nothing here is scored.
sft-prompts/train and rl-prompts/train: the same prompt pool, deduplicated across sources with responses
merged and HH/PKU test prompts removed. rl-prompts additionally marks selection=pku_label_conflict where PKU's
better and safer labels disagree with opposite safety flags; preference_pairs indexes those responses.
This is an annotation, not a… See the full description on the dataset page: https://huggingface.co/datasets/RLLab/safe-alignment-dynamic.SpeechTextMatching_Tedlium2Train
Dataset Card for "SpeechTextMatching_TEDLIUM2Train"
More Information needed
EnhancementDetection_LibrittsTrainClean360Wham
Dataset Card for "EnhancementDetection_LibrittsTrainClean360Wham"
More Information needed
SpeakerVerification_Aishell1Train
Dataset Card for "SpeakerVerification_AISHELL1Train"
More Information needed
DynamicvlmSpeechTextMatching_LibrispeechTrainClean360
Dataset Card for "speechTextMatching_LibrispeechTrainClean360"
More Information needed
Disney-Theme-Park-Queue-Dynamics
🎢 Disney World Queue Dynamics
A Comprehensive EDA & Strategic Analysis
Author: Matan Zigelman • University: Reichman University • Date: March 2026
📋 Project Introduction: Disney Theme Park Queue Dynamics
This project analyzes a numeric-heavy operational dataset from a major Disney theme park, sourced from kaggle and containing approximately 3,757,301 records. The dataset is primarily driven by time-based and operational metrics ( e.g.… See the full description on the dataset page: https://huggingface.co/datasets/matanzig/Disney-Theme-Park-Queue-Dynamics.SpeechDetection_Aishell1Train
Dataset Card for "SpeechDetection_AISHELL1Train"
More Information needed
protein_dynamic_properties
Protein Sequences and Dynamic Properties for Training SeqDance and ESMDance
This dataset contains protein sequences, dynamic properties, and feature weights used for training SeqDance and ESMDance, two protein language models designed to learn protein dynamic properties.
training_test_data_sequence.csv
This file contains 64,403 protein sequences with associated metadata for training and testing (excluding dynamicPDB).
Columns:
name – Unique identifier for each… See the full description on the dataset page: https://huggingface.co/datasets/ChaoHou/protein_dynamic_properties.ChordClassification_AcousticGuitarAndPiano
Dataset Card for "chord_classification_acoustic_guitar_and_piano"
More Information needed
NoiseSNRLevelPredictionGaussian_VoxcelebMusan
Dataset Card for "NoiseSNRLevelPredictiongaussian_VoxcelebMusan"
More Information needed
SpeakerVerification_Tedlium2Train
Dataset Card for "SpeakerVerification_TEDLIUM2Train"
More Information needed
NoiseSNRLevelPredictionNoise_VoxcelebMusan
Dataset Card for "NoiseSNRLevelPredictionnoise_VoxcelebMusan"
More Information needed
SpeechDetection_Tedlium2Train
Dataset Card for "speechDetection_TEDLIUM2Train"
More Information needed
SpokenTermDetection_Tedlium2Train
Dataset Card for "SpokenTermDetection_Tedlium2Train"
More Information needed
fresh-swe-pro-dynamic
Dataset Card
Dataset Description
Fresh SWE-Pro Dynamic is a source-verified, Docker-executable benchmark for software-engineering agents. The public release contains five recent repository tasks with pinned base commits, problem statements, gold patches, regression tests, Docker images, source provenance, and validation evidence.
Task: software-engineering agent evaluation and patch generation
Language: English
Public size: 5 instances
Release: 2026.09
Source… See the full description on the dataset page: https://huggingface.co/datasets/LianeMarilin/fresh-swe-pro-dynamic.generalization-dynamics-evals
Generalization Dynamics — Main Eval Suite
Prepared test sets for the 6 main evaluation families from
Generalization dynamics across fine-tuning
(Table 1).
Use with the unified runner:
https://github.com/jiaxin-wen/FT-generalization/tree/main/release
from huggingface_hub import snapshot_download
root = snapshot_download(
repo_id="jiaxin-wen/generalization-dynamics-evals", repo_type="dataset")
Or browse a single task (the dataset viewer shows all configs):
from datasets… See the full description on the dataset page: https://huggingface.co/datasets/jiaxin-wen/generalization-dynamics-evals.DynamicReplica_wai
DynamicReplica Dataset in WAI format
Preprocessed DynamicReplica following MapAnything.
Each scene contains the following structure when extracted:
0a5b4c-3_obj_source/
├── depth
| ├── 0a5b4c-3_obj_source_left-0000.exr
| ├── ...
├── images
| ├── 0a5b4c-3_obj_source_left-0000.png
| ├── ...
├── _process_log_backup.json
├── _process_log.json
└── scene_meta.json
Data Format Details:
depth: depth in .exr format.
images: images in .png format.
scene_meta.json: meta… See the full description on the dataset page: https://huggingface.co/datasets/ZhengGuangze/DynamicReplica_wai.MARBLEGenreClassification_MTG-Genre-Fold1
Dataset Card for "MARBLEGenreClassification_MTG-Genre-Fold1"
More Information needed
update-interrupt-benchmark
Update-Driven Math & Code Interrupt Datasets
Paper: Are Large Reasoning Models Interruptible?
Authors: Tsung-Han Wu*, Mihran Miroyan*, David Chan, Trevor Darrell, Narges Norouzi, Joseph Gonzalez
Project page: https://dynamic-lm.github.io/
Github: https://github.com/dynamic-lm/interrupt-lrm
This dataset page contains the update-driven interrupt subsets for math (GSM8K, MATH500, AIME) and coding (LiveCodeBench) problems. For both splits, we revise the source problems and… See the full description on the dataset page: https://huggingface.co/datasets/dynamic-lm/update-interrupt-benchmark.dynamic_earthnet
GeoBench-2 Dataset License Attribution
Dataset Name: m-DynamicEarthNetOriginal Dataset Name: DynamicEarthNetOriginal Source: https://mediatum.ub.tum.de/1650201
Related Publication(s): https://doi.org/10.1109/CVPR52688.2022.02048
Licensing
Annotation License: CC BY-SA 4.0
Image License: Planet Labs “Planet Fusion” imagery — licence terms as provided by the dataset host (via Mediatum) under BY-SA.
Declared By Original Provider: https://mediatum.ub.tum.de/1650201… See the full description on the dataset page: https://huggingface.co/datasets/aialliance/dynamic_earthnet.
