CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01AlphaDojo /dojo_market_dynamicstext1K<n<10K0 likes12k downloads1mo agoHugging Face02Shivamkak /STUZero-Atari-Dynamics STUZero Atari Dynamics Dataset Offline dynamics training datasets collected from trained EfficientZero V2 (EZv2) benchmark models on Atari games. Each game's data is stored in a subfolder named {game}_{steps} indicating the game and the number of training steps of the source checkpoint. While all models were trained for 120K steps, best results in some games were attained at earlier checkpoints. The model with best eval scores was used to curate data for each game.… See the full description on the dataset page: https://huggingface.co/datasets/Shivamkak/STUZero-Atari-Dynamics.textreinforcement-learningn<1K0 likes4.8k downloads6mo agoHugging Face03DynamicIntelligence /humanoid-robots-training-dataset Dynamic Intelligence — Humanoid Robot Training Dataset A first-person (egocentric) video dataset of human hand manipulation, designed for training humanoid robot policies via imitation learning. Each episode captures a person performing an everyday household task — folding clothes, moving dishes, opening doors — filmed from a head-mounted iPhone using its built-in LiDAR and depth sensors. The dataset pairs each video with frame-level 3D hand tracking and camera pose data, giving… See the full description on the dataset page: https://huggingface.co/datasets/DynamicIntelligence/humanoid-robots-training-dataset.tabularrobotics10K<n<100K0 likes2.6k downloads6mo agoHugging Face04EverMind-AI /EverMemBench-Dynamic EverMemBench-Dynamic A benchmark dataset for evaluating long-term memory capabilities in conversational AI systems. It is part of EverMemBench, the first benchmark designed for long-horizon collaborative memory, introduced in the paper Evaluating Long-Horizon Memory for Multi-Party Collaborative Dialogues — accepted at KDD 2026 (Oral). Configurations This dataset has three configurations (subsets): dialogues Multi-turn group dialogues spanning ~250… See the full description on the dataset page: https://huggingface.co/datasets/EverMind-AI/EverMemBench-Dynamic.textquestion-answering1K<n<10K5 likes952 downloads2mo agoHugging Face05dynamicfeed /live-facts-snapshot Live Facts Snapshot A daily snapshot of verifiable, post-training-cutoff world-state facts — the kind of ground truth language models cannot know from training data — exported through Dynamic Feed, a live, verifiable data API whose every response is Ed25519-signed. One file per day (data/YYYY-MM-DD.jsonl), one fact per line, and every row carries its own source, source_url and measured_at. Facts covered per day: tool facts upstream source licence software_version… See the full description on the dataset page: https://huggingface.co/datasets/dynamicfeed/live-facts-snapshot.textquestion-answering1K<n<10K0 likes759 downloads8h agoHugging Face06TokenWasteGroup /DynamicMCPBench DynamicMCPBench A trace-grounded, effect-scored benchmark for LLM agents on live MCP servers. Tasks are generated forward: an explorer agent drives real MCP tools until a goal is reached, the recorded trace is distilled into a TaskSpec, and candidates are graded on whether they reproduce the effects the trace produced — checkpoints, equivalence sets, minefields, a partial order — never on matching an answer string or a fixed tool list. Candidates are evaluated under… See the full description on the dataset page: https://huggingface.co/datasets/TokenWasteGroup/DynamicMCPBench.textother1K<n<10K1 likes648 downloads25d agoHugging Face07torchgeo /dynamic_earthnetDynamic EarthNet dataset redistributed from https://mediatum.ub.tum.de/1650201 and https://cvg.cit.tum.de/webshare/u/toker/dynnet_training_splits/ under a common tarball for simpler download speeds. Individual zip files were replaced with tarballs instead. In the mediatum server version the following directories have the wrong name compared to the given split txt files: /labels/5111_4560_13_38S /labels/6204_3495_13_46N /labels/7026_3201_13_52N /labels/7367_5050_13_54S /labels/2459_4406_13_19S… See the full description on the dataset page: https://huggingface.co/datasets/torchgeo/dynamic_earthnet.image1M<n<10M2 likes574 downloads2y agoHugging Face08mthreetw /Semantic-Flow-Dynamics-SFD Semantic Flow Dynamics (SFD) — A Formally Specified Social-Science Theory Corpus TL;DR: 614 Chinese-language formalized social-science concepts across 25 papers, UUID-linked with typed derivation relations (derives_from, leads_to, falsified_by, …) — usable for knowledge-graph construction, RAG over structured theory, or as a Chinese formal-reasoning corpus. Author: 黃正宇 Cheng Yu HuangContact: mthree.tw@gmail.com What This Dataset Is This corpus is an ongoing… See the full description on the dataset page: https://huggingface.co/datasets/mthreetw/Semantic-Flow-Dynamics-SFD.textgraph-ml1K<n<10K0 likes496 downloads10d agoHugging Face09EigenformAI /groundtruth-dynamic-benchmarking Groundtruth Dynamic Benchmarking — Geology Question sets and grading rubrics for evaluating LLMs on real-world geological reasoning. Every question is authored from a real source corpus, and every claim in the grading key carries an evidence locator back to that corpus — nothing is synthetic. Licensing/redistribution status varies by corpus — see License. This dataset holds the questions, grading rubrics, and source corpora. Running an evaluation (generating answers from a model… See the full description on the dataset page: https://huggingface.co/datasets/EigenformAI/groundtruth-dynamic-benchmarking.textquestion-answeringn<1K0 likes387 downloads26d agoHugging Face10RLLab /safe-alignment-dynamic safe-alignment-dynamic Training prompts for score-conditioned SFT / RL and separate reward-model pair sets; nothing here is scored. sft-prompts/train and rl-prompts/train: the same prompt pool, deduplicated across sources with responses merged and HH/PKU test prompts removed. rl-prompts additionally marks selection=pku_label_conflict where PKU's better and safer labels disagree with opposite safety flags; preference_pairs indexes those responses. This is an annotation, not a… See the full description on the dataset page: https://huggingface.co/datasets/RLLab/safe-alignment-dynamic.tabular100K<n<1M0 likes365 downloads11d agoHugging Face11DynamicSuperbPrivate /SpeechTextMatching_Tedlium2Train Dataset Card for "SpeechTextMatching_TEDLIUM2Train" More Information needed audio10K<n<100K0 likes321 downloads3y agoHugging Face12DynamicSuperbPrivate /EnhancementDetection_LibrittsTrainClean360Wham Dataset Card for "EnhancementDetection_LibrittsTrainClean360Wham" More Information needed audio100K<n<1M0 likes317 downloads3y agoHugging Face13DynamicSuperbPrivate /SpeakerVerification_Aishell1Train Dataset Card for "SpeakerVerification_AISHELL1Train" More Information needed audio100K<n<1M0 likes299 downloads3y agoHugging Face14yu2hi13 /Dynamicvlmimage100K<n<1M0 likes266 downloads1y agoHugging Face15DynamicSuperbPrivate /SpeechTextMatching_LibrispeechTrainClean360 Dataset Card for "speechTextMatching_LibrispeechTrainClean360" More Information needed audio100K<n<1M0 likes259 downloads3y agoHugging Face16matanzig /Disney-Theme-Park-Queue-Dynamics 🎢 Disney World Queue Dynamics A Comprehensive EDA & Strategic Analysis Author: Matan Zigelman • University: Reichman University • Date: March 2026 📋 Project Introduction: Disney Theme Park Queue Dynamics This project analyzes a numeric-heavy operational dataset from a major Disney theme park, sourced from kaggle and containing approximately 3,757,301 records. The dataset is primarily driven by time-based and operational metrics ( e.g.… See the full description on the dataset page: https://huggingface.co/datasets/matanzig/Disney-Theme-Park-Queue-Dynamics.imagetabular-classification1M<n<10M2 likes243 downloads6mo agoHugging Face17DynamicSuperbPrivate /SpeechDetection_Aishell1Train Dataset Card for "SpeechDetection_AISHELL1Train" More Information needed audio100K<n<1M0 likes235 downloads3y agoHugging Face18ChaoHou /protein_dynamic_properties Protein Sequences and Dynamic Properties for Training SeqDance and ESMDance This dataset contains protein sequences, dynamic properties, and feature weights used for training SeqDance and ESMDance, two protein language models designed to learn protein dynamic properties. training_test_data_sequence.csv This file contains 64,403 protein sequences with associated metadata for training and testing (excluding dynamicPDB). Columns: name – Unique identifier for each… See the full description on the dataset page: https://huggingface.co/datasets/ChaoHou/protein_dynamic_properties.text100K<n<1M2 likes230 downloads8mo agoHugging Face19DynamicSuperb /ChordClassification_AcousticGuitarAndPiano Dataset Card for "chord_classification_acoustic_guitar_and_piano" More Information needed audion<1K1 likes207 downloads3y agoHugging Face20DynamicSuperbPrivate /NoiseSNRLevelPredictionGaussian_VoxcelebMusan Dataset Card for "NoiseSNRLevelPredictiongaussian_VoxcelebMusan" More Information needed audio10K<n<100K0 likes206 downloads3y agoHugging Face21DynamicSuperbPrivate /SpeakerVerification_Tedlium2Train Dataset Card for "SpeakerVerification_TEDLIUM2Train" More Information needed audio10K<n<100K0 likes200 downloads3y agoHugging Face22DynamicSuperbPrivate /NoiseSNRLevelPredictionNoise_VoxcelebMusan Dataset Card for "NoiseSNRLevelPredictionnoise_VoxcelebMusan" More Information needed audio10K<n<100K0 likes199 downloads3y agoHugging Face23DynamicSuperbPrivate /SpeechDetection_Tedlium2Train Dataset Card for "speechDetection_TEDLIUM2Train" More Information needed audio10K<n<100K0 likes198 downloads3y agoHugging Face24DynamicSuperbPrivate /SpokenTermDetection_Tedlium2Train Dataset Card for "SpokenTermDetection_Tedlium2Train" More Information needed audio10K<n<100K0 likes194 downloads3y agoHugging Face25LianeMarilin /fresh-swe-pro-dynamic Dataset Card Dataset Description Fresh SWE-Pro Dynamic is a source-verified, Docker-executable benchmark for software-engineering agents. The public release contains five recent repository tasks with pinned base commits, problem statements, gold patches, regression tests, Docker images, source provenance, and validation evidence. Task: software-engineering agent evaluation and patch generation Language: English Public size: 5 instances Release: 2026.09 Source… See the full description on the dataset page: https://huggingface.co/datasets/LianeMarilin/fresh-swe-pro-dynamic.texttext-generationn<1K0 likes187 downloads19d agoHugging Face26jiaxin-wen /generalization-dynamics-evals Generalization Dynamics — Main Eval Suite Prepared test sets for the 6 main evaluation families from Generalization dynamics across fine-tuning (Table 1). Use with the unified runner: https://github.com/jiaxin-wen/FT-generalization/tree/main/release from huggingface_hub import snapshot_download root = snapshot_download( repo_id="jiaxin-wen/generalization-dynamics-evals", repo_type="dataset") Or browse a single task (the dataset viewer shows all configs): from datasets… See the full description on the dataset page: https://huggingface.co/datasets/jiaxin-wen/generalization-dynamics-evals.texttext-classification10K<n<100K0 likes183 downloads4mo agoHugging Face27ZhengGuangze /DynamicReplica_wai DynamicReplica Dataset in WAI format Preprocessed DynamicReplica following MapAnything. Each scene contains the following structure when extracted: 0a5b4c-3_obj_source/ ├── depth | ├── 0a5b4c-3_obj_source_left-0000.exr | ├── ... ├── images | ├── 0a5b4c-3_obj_source_left-0000.png | ├── ... ├── _process_log_backup.json ├── _process_log.json └── scene_meta.json Data Format Details: depth: depth in .exr format. images: images in .png format. scene_meta.json: meta… See the full description on the dataset page: https://huggingface.co/datasets/ZhengGuangze/DynamicReplica_wai.image100K<n<1M0 likes172 downloads11mo agoHugging Face28DynamicSuperb /MARBLEGenreClassification_MTG-Genre-Fold1 Dataset Card for "MARBLEGenreClassification_MTG-Genre-Fold1" More Information needed audion<1K0 likes169 downloads2y agoHugging Face29dynamic-lm /update-interrupt-benchmark Update-Driven Math & Code Interrupt Datasets Paper: Are Large Reasoning Models Interruptible? Authors: Tsung-Han Wu*, Mihran Miroyan*, David Chan, Trevor Darrell, Narges Norouzi, Joseph Gonzalez Project page: https://dynamic-lm.github.io/ Github: https://github.com/dynamic-lm/interrupt-lrm This dataset page contains the update-driven interrupt subsets for math (GSM8K, MATH500, AIME) and coding (LiveCodeBench) problems. For both splits, we revise the source problems and… See the full description on the dataset page: https://huggingface.co/datasets/dynamic-lm/update-interrupt-benchmark.texttext-generation1K<n<10K4 likes165 downloads11mo agoHugging Face30aialliance /dynamic_earthnet GeoBench-2 Dataset License Attribution Dataset Name: m-DynamicEarthNetOriginal Dataset Name: DynamicEarthNetOriginal Source: https://mediatum.ub.tum.de/1650201 Related Publication(s): https://doi.org/10.1109/CVPR52688.2022.02048 Licensing Annotation License: CC BY-SA 4.0 Image License: Planet Labs “Planet Fusion” imagery — licence terms as provided by the dataset host (via Mediatum) under BY-SA. Declared By Original Provider: https://mediatum.ub.tum.de/1650201… See the full description on the dataset page: https://huggingface.co/datasets/aialliance/dynamic_earthnet.textn<1K0 likes154 downloads10mo agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.