CoolFace
30 shown

datasets

Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.

Clear all
01twelvedata /financial-world-model Twelve Data World Model Dataset A multi-modal financial time-series dataset built from Twelve Data market data. Each timeframe is published in three parallel views: bars_* — OHLCV bars enriched with causal technical indicators and macro context, in Parquet. text_* — instruction-tuning prompts/labels derived from the bars, in JSONL. trajectories_* — fixed-length rolling windows of state vectors plus next-state pairs, suitable for world-model / sequence-model training, in… See the full description on the dataset page: https://huggingface.co/datasets/twelvedata/financial-world-model.tabulartime-series-forecasting10M<n<100M8 likes2.3k downloads6h agoHugging Face02Rapidata /world-model-physics Rapidata Physics Benchmark Built by Rapidata. Do video and world models understand physics? We gave 25 video- and world models the same real-world starting frame and scene description from Physics-IQ and asked them to predict what happens next. ~283,000 human votes, collected with the Rapidata Python SDK, decided which continuation is more realistic — with the real recording competing as a hidden 26th participant. Each row is a head-to-head matchup between two participants on… See the full description on the dataset page: https://huggingface.co/datasets/Rapidata/world-model-physics.tabulartext-to-video10K<n<100K0 likes470 downloads6d agoHugging Face03Ryukijano /repro-causal-jepa-learning-world-models-through-object-level-latent-masking-traces Agent traces Agent sessions published from a Trackio Logbook. tabularn<1K0 likes91 downloads2mo agoHugging Face04thuml /webarena-world-model-cotSee https://github.com/thuml/RLVR-World for examples for using this dataset. Citation @article{wu2025rlvr, title={RLVR-World: Training World Models with Reinforcement Learning}, author={Jialong Wu and Shaofeng Yin and Ningya Feng and Mingsheng Long}, journal={arXiv preprint arXiv:2505.13934}, year={2025}, } tabular1K<n<10K0 likes77 downloads1y agoHugging Face05yilin-wu /world_model_real_rollout_gentabular1K<n<10K0 likes77 downloads5mo agoHugging Face06LangAGI-Lab /world_model_for_wa_desc_with_tao_dataset Dataset Card for "world_model_for_wa_desc_with_tao_dataset" More Information needed tabular10K<n<100K0 likes69 downloads2y agoHugging Face07Clementppr /lerobot_pick_and_place_dataset_world_modelThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v2.1", "robot_type": "so100", "total_episodes": 30, "total_frames": 13572, "total_tasks": 1, "total_videos": 30, "total_chunks": 1, "chunks_size": 1000, "fps": 30, "splits": { "train": "0:30"}, "data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Clementppr/lerobot_pick_and_place_dataset_world_model.tabularrobotics10K<n<100K0 likes58 downloads1y agoHugging Face08rl26-world-models /so101-task2-720p-whole-arm-v3-cleanedThis dataset was created using LeRobot. Dataset Description Recovered and cleaned SO-101 task 2 dataset. Bad final source episodes 95 and 96 were removed. V3 data, episode metadata, and video shard indices are compact and contiguous. License: apache-2.0 Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so_follower", "total_episodes": 95, "total_frames": 73988, "total_tasks": 1, "chunks_size": 1000… See the full description on the dataset page: https://huggingface.co/datasets/rl26-world-models/so101-task2-720p-whole-arm-v3-cleaned.tabularrobotics10K<n<100K0 likes49 downloads4mo agoHugging Face09ZaidGhazal /world-models-eval DreamGrasp: Processed LIBERO Manipulation Demonstrations Does a robot policy's evaluation still mean something if it never touched a real simulator, only a world model's imagination of one? This dataset is the shared training data behind that question, a single, ready-to-train release built from LIBERO's manipulation demonstrations (libero_spatial, libero_object, libero_goal). It provides: Fixed, versioned train / validation / test / held-out splits, so every result trained on… See the full description on the dataset page: https://huggingface.co/datasets/ZaidGhazal/world-models-eval.tabularrobotics100K<n<1M0 likes46 downloads3mo agoHugging Face10Ryukijano /marl-world-model-lerobottabularn<1K0 likes37 downloads1mo agoHugging Face11drvp /world_model_eval_logstabularn<1K0 likes30 downloads2mo agoHugging Face12rl26-world-models /so101-task2-720p-whole-arm-v4-fresh-trimmedThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so_follower", "total_episodes": 50, "total_frames": 38664, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:50" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/rl26-world-models/so101-task2-720p-whole-arm-v4-fresh-trimmed.tabularrobotics10K<n<100K0 likes24 downloads4mo agoHugging Face13rl26-world-models /so101-task2-720p-whole-arm-v4-freshThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so_follower", "total_episodes": 50, "total_frames": 38664, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:50" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/rl26-world-models/so101-task2-720p-whole-arm-v4-fresh.tabularrobotics10K<n<100K1 likes23 downloads4mo agoHugging Face14rl26-world-models /so101-task2-720p-whole-arm-v3-cleaned-trimmedThis dataset was created using LeRobot. Dataset Description Recovered and cleaned SO-101 task 2 dataset. Bad final source episodes 95 and 96 were removed. V3 data, episode metadata, and video shard indices are compact and contiguous. License: apache-2.0 Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so_follower", "total_episodes": 95, "total_frames": 73988, "total_tasks": 1, "chunks_size": 1000… See the full description on the dataset page: https://huggingface.co/datasets/rl26-world-models/so101-task2-720p-whole-arm-v3-cleaned-trimmed.tabularrobotics10K<n<100K0 likes22 downloads4mo agoHugging Face15LangAGI-Lab /world_model_for_wa_acctree_dataset_14K Dataset Card for "world_model_for_wa_acctree_dataset_14K" More Information needed tabular10K<n<100K1 likes20 downloads2y agoHugging Face16rl26-world-models /inference-viz-1This dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so101_follower", "total_episodes": 1, "total_frames": 60, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:1" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/rl26-world-models/inference-viz-1.tabularroboticsn<1K0 likes19 downloads4mo agoHugging Face17ultrastar111 /sokoban_easy_v8_noncot_chunk_k10_world_model_20260622_perseg sokoban_easy_v8_noncot_chunk_k10_world_model_20260622_perseg Sokoban action-conditioned visual world-model SFT data (non-CoT baseline) for the BAGEL-7B-MoT VLM-Gym feedback-interval study. Format: gzipped JSONL shards under training/, one packed row = one episode. Frames are base64 JPEG (q95). Per-segment CoT layout: <think> per-step imagined frame (MSE target) </think> then the committed action chunk; between chunks a loss-0 "Action executed." + real frame (GT re-grounding).… See the full description on the dataset page: https://huggingface.co/datasets/ultrastar111/sokoban_easy_v8_noncot_chunk_k10_world_model_20260622_perseg.tabularreinforcement-learning100K<n<1M0 likes17 downloads2mo agoHugging Face18julie-trrsr /SO101-world-model-5fpsThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so_follower", "total_episodes": 1, "total_frames": 107, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 5, "splits": { "train": "0:1" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/julie-trrsr/SO101-world-model-5fps.tabularroboticsn<1K0 likes16 downloads5mo agoHugging Face19rl26-world-models /so101-task1-720p-whole-arm-trimmed-subsampled-10fpsThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0","robot_type": "so_follower", "total_episodes": 201, "total_frames": 62169, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100,"video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:201" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/rl26-world-models/so101-task1-720p-whole-arm-trimmed-subsampled-10fps.tabularrobotics10K<n<100K0 likes16 downloads4mo agoHugging Face20rl26-world-models /so101-task1-720p-whole-arm-trimmedThis dataset was created using LeRobot. Dataset Structure meta/info.json: { "codebase_version": "v3.0", "robot_type": "so_follower", "total_episodes": 201, "total_frames": 62169, "total_tasks": 1, "chunks_size": 1000, "data_files_size_in_mb": 100, "video_files_size_in_mb": 200, "fps": 30, "splits": { "train": "0:201" }, "data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet", "video_path":… See the full description on the dataset page: https://huggingface.co/datasets/rl26-world-models/so101-task1-720p-whole-arm-trimmed.tabularrobotics10K<n<100K0 likes16 downloads4mo agoHugging Face21ultrastar111 /sokoban_easy_v8_noncot_chunk_k5_world_model_20260622_perseg sokoban_easy_v8_noncot_chunk_k5_world_model_20260622_perseg Sokoban action-conditioned visual world-model SFT data (non-CoT baseline) for the BAGEL-7B-MoT VLM-Gym feedback-interval study. Format: gzipped JSONL shards under training/, one packed row = one episode. Frames are base64 JPEG (q95). Per-segment CoT layout: <think> per-step imagined frame (MSE target) </think> then the committed action chunk; between chunks a loss-0 "Action executed." + real frame (GT re-grounding).… See the full description on the dataset page: https://huggingface.co/datasets/ultrastar111/sokoban_easy_v8_noncot_chunk_k5_world_model_20260622_perseg.tabularreinforcement-learning100K<n<1M0 likes15 downloads2mo agoHugging Face22rl26-world-models /so101-task2-720p-whole-arm-v3-cleaned-trimmed-subsampled-10fpstabular10K<n<100K0 likes14 downloads4mo agoHugging Face23rl26-world-models /so101-task2-720p-whole-arm-cube-trimmedtabular10K<n<100K0 likes14 downloads4mo agoHugging Face24ultrastar111 /sokoban_easy_v8_cot_chunk_kinf_world_model_20260707_perseg sokoban_easy_v8_cot_chunk_kinf_world_model_20260707_perseg Sokoban action-conditioned visual world-model SFT data (CoT self-rollout) for the BAGEL-7B-MoT VLM-Gym feedback-interval study. Format: gzipped JSONL shards under training/, one packed row = one episode. Frames are base64 JPEG (q95). Per-segment CoT layout: <think> per-step imagined frame (MSE target) </think> then the committed action chunk; between chunks a loss-0 "Action executed." + real frame (GT re-grounding). See… See the full description on the dataset page: https://huggingface.co/datasets/ultrastar111/sokoban_easy_v8_cot_chunk_kinf_world_model_20260707_perseg.tabularreinforcement-learning100K<n<1M0 likes14 downloads2mo agoHugging Face25ultrastar111 /sokoban_easy_v8_noncot_chunk_k3_world_model_20260622_perseg sokoban_easy_v8_noncot_chunk_k3_world_model_20260622_perseg Sokoban action-conditioned visual world-model SFT data (non-CoT baseline) for the BAGEL-7B-MoT VLM-Gym feedback-interval study. Format: gzipped JSONL shards under training/, one packed row = one episode. Frames are base64 JPEG (q95). Per-segment CoT layout: <think> per-step imagined frame (MSE target) </think> then the committed action chunk; between chunks a loss-0 "Action executed." + real frame (GT re-grounding).… See the full description on the dataset page: https://huggingface.co/datasets/ultrastar111/sokoban_easy_v8_noncot_chunk_k3_world_model_20260622_perseg.tabularreinforcement-learning100K<n<1M0 likes13 downloads2mo agoHugging Face26rl26-world-models /so101-task1-720p-whole-arm-cube-trimmedtabular10K<n<100K0 likes10 downloads4mo agoHugging Face27LangAGI-Lab /world_model_for_wa_train_dataset Dataset Card for "world_model_for_wa_train_dataset" More Information needed tabular10K<n<100K0 likes9 downloads2y agoHugging Face28LangAGI-Lab /world_model_for_wa_tao_datasettabular10K<n<100K0 likes9 downloads2y agoHugging Face29LangAGI-Lab /world_model_for_wa_qa_datasettabular10K<n<100K0 likes9 downloads2y agoHugging Face30LangAGI-Lab /world_model_for_wa_desc_with_tao_dataset_with_transition_counttabular10K<n<100K0 likes9 downloads2y agoHugging Face

Listings come live from the Hugging Face Hub API. CoolFace does not host these files.