datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
action-atlas-groot-activationsactionnet-subset100-gtdepth
ActionNet subset100 with ground-truth depth
This is a 100-episode subset redistribution of a third-party dataset, plus our derived
artifacts. Read the attribution before using it.
Attribution and license
Upstream dataset
FourierIntelligence/ActionNet (Fourier Intelligence)
What is redistributed
100 episodes out of 30,121, byte-identical to the upstream tars: rgb.mp4 (1280x800 fisheye), depth.mkv (lossless 16-bit), timestamps.json, <ULID>.hdf5… See the full description on the dataset page: https://huggingface.co/datasets/glory-hyeok/actionnet-subset100-gtdepth.action-atlas-oft-activationsgithub-actionsactionWM_vfx_sample150
H3 VFX editing: 150 review examples
150 synthetic VFX event videos for colleague review, generated with MiniMax H3 Max
text-to-video and balanced prompt expansion. Each raw MP4 and its exact submitted
raw prompt TXT share a task ID and are stored together in this directory.
The text column in the dataset viewer contains the same raw prompt.
Category
First person
Third person
Total
Global environment
30
30
60
Local object
30
30
60
Character effects
0
30
30
Total… See the full description on the dataset page: https://huggingface.co/datasets/Geral-Yuan/actionWM_vfx_sample150.jma-gsi-disaster-action-corpus
JMA-GSI Disaster Action Corpus
A grounded, multilingual disaster-response dataset built from official Japanese government open data (JMA alert XML + JMA multilingual glossary + JMA forecast-area GIS + GSI designated evacuation shelters). Structured hazard alerts are transformed into easy-Japanese and multilingual (ja / easy-ja / en / vi / id / ne / my) action guidance, linked to hazard-compatible evacuation shelters, with full source traceability.
License (derived dataset): CC BY… See the full description on the dataset page: https://huggingface.co/datasets/edomaru/jma-gsi-disaster-action-corpus.gpio-llm-rpi5-actions
GPIO-LLM: Raspberry Pi 5 GPIO request-to-action dataset
Requests to a Raspberry Pi 5 in plain English, paired with the structured, validated GPIO action a
small on-device model should produce: a hardware operation, a clarifying question when the pin or device
is unknown, or a refusal when the request is invalid or unsafe. It was built to train a ~20M-parameter
English model that runs offline on the Pi.
Safety. Model output must never drive hardware directly. Every action is… See the full description on the dataset page: https://huggingface.co/datasets/AwaleSagar/gpio-llm-rpi5-actions.corporate-actions
US Corporate Actions — dividends and splits
391 639 dividends from 3 327 filers · 5 619 splits from 3 814 filers ·
2005 to 2026
Built to close a specific hole. A filing states shares and earnings per share
as of the day it was made; every price series is adjusted for splits since.
Multiply one by the other and the answer is wrong by the split factor — on
Deckers that turned a 6.9% earnings yield into 41.7%, a P/E of 1.8.
The pipeline lives in recipe/ at the same revision as the… See the full description on the dataset page: https://huggingface.co/datasets/ZipLime/corporate-actions.action15s-media-20260910Media and bilingual annotations for the companion gameplay action review.
Use train_15s.jsonl as the current accepted selection: each row contains its relative video path, checksum, source interval, and English/Chinese timed action labels. Historical media files from earlier progress snapshots may remain in the repository; only the manifest defines the current batch. summary.json reports coverage and pending visual screening separately. The default dataset configuration reads only the training… See the full description on the dataset page: https://huggingface.co/datasets/mikusama99/action15s-media-20260910.tb2-eval-qwen3-8b-action-clean-6ep
Terminal-Bench 2.0 eval results - Qwen3-8B action-clean 6ep SFT
Compact Harbor eval results for violetxi/qwen3-8b-terminal-action-clean-6ep on
Terminal-Bench 2.0 using the terminus-2 agent harness.
Each dataset split is one checkpoint step. Rows are Harbor trial directories and include binary reward,
per-test-case pass/fail data from verifier/ctrf.json, exception text when present, and run metadata.
Raw terminal recordings, panes, and completion logs are not included.… See the full description on the dataset page: https://huggingface.co/datasets/violetxi/tb2-eval-qwen3-8b-action-clean-6ep.libero_90_various_action_spacesHS_gello_260813-action_is_ur5state_cam1_remappi0_stacking_action_chunkscraft-multiturn-actions-split-nothinkstack_absolute_actions_v4This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": null,
"total_episodes": 51,
"total_frames": 12244,
"total_tasks": 1,
"total_videos": 153,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 15,
"splits": {
"train": "0:51"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/danielsanjosepro/stack_absolute_actions_v4.act_stacking_action_chunksGr00t_lerobot_state_action_all_bottom_boxes_22april_12pmThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "xarm",
"total_episodes": 899,
"total_frames": 263674,
"total_tasks": 5,
"total_videos": 2697,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:899"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Anas0711/Gr00t_lerobot_state_action_all_bottom_boxes_22april_12pm.cs2-action-inference-test
CS2 战术 Action 推理测试集
本测试集用于 WAN I2V 的战术动作定性测试。每个小类只保留 1 张真实比赛 POV 第一帧,以及两种英文文本条件;本版不提供 GT 视频。第一帧来源依据 parse-dem 的 events.csv、game_events.csv 或逐 tick 状态对齐到 opencs2_matches* 视频。
数据约定
共 45 个 case、9 个大类。
每个 case 只有一张 832x480 的 first_frame.png,作为 WAN I2V 条件图;不裁剪或复制 GT clip。首帧优先选择正常持械、水平视角、无遮挡且较开阔的画面。
prompt.txt 是完整英文 prompt,包含首帧可见环境、初始持械状态、画面保持要求和整段唯一动作变化。
chunk_prompts.json 固定包含 5 个英文 prompt,依次描述期望生成视频的 0-1、1-2、2-3、3-4、4-5 秒。
metadata.json… See the full description on the dataset page: https://huggingface.co/datasets/mikusama99/cs2-action-inference-test.JL-ActionBoundary-1K-v1.0.0
JL-ActionBoundary-1K v1.0.0
Counterfactual Ask–Inspect–Act–Defer supervision for coding agents
JL-ActionBoundary-1K teaches a coding agent to choose the correct next policy before changing code:
ACT: the task is sufficiently specified for bounded repository work;
INSPECT: missing information can be recovered from the repository;
ASK: a material product decision belongs to the user;
DEFER: live execution authority or rollback ownership is missing.… See the full description on the dataset page: https://huggingface.co/datasets/jumplander/JL-ActionBoundary-1K-v1.0.0.stack_cake_v2_absolute_actionsThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": null,
"total_episodes": 101,
"total_frames": 23884,
"total_tasks": 1,
"total_videos": 303,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 15,
"splits": {
"train": "0:101"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/LSY-lab/stack_cake_v2_absolute_actions.libero_goal_force_openvla_predicted_action_fixedactionWM_vfx_h3
ActionWM VFX — 600 H3 videos
600 synthetic VFX videos generated with MiniMax H3 Max text-to-video and balanced
prompt expansion. Each unchanged original videos/TASK_ID.mp4 matches
prompts/TASK_ID.txt. The TXT and viewer text column contain the exact raw prompt
submitted before the model's server-side prompt expansion.
Category
First person
Third person
Total
Global environment
120
120
240
Local object
120
120
240
Character effects
0
120
120
Total
240
360
600… See the full description on the dataset page: https://huggingface.co/datasets/Geral-Yuan/actionWM_vfx_h3.0721_change_actionminiwob_actions_onhot
Dataset Card for "miniwob_actions_onhot"
More Information needed
astra_grab_floor_toys_without_observations_actionsThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.0",
"robot_type": "astra",
"total_episodes": 50,
"total_frames": 73944,
"total_tasks": 1,
"total_videos": 150,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:50"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/lookas/astra_grab_floor_toys_without_observations_actions.libero_spatial_various_action_spaces_no_noopsThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "franka",
"total_episodes": 439,
"total_frames": 53889,
"total_tasks": 10,
"total_videos": 878,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:439"},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/sean1295/libero_spatial_various_action_spaces_no_noops.drawer_absolute_actions_v2This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": null,
"total_episodes": 75,
"total_frames": 25156,
"total_tasks": 1,
"total_videos": 225,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 15,
"splits": {
"train": "0:75"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/danielsanjosepro/drawer_absolute_actions_v2.orange_bobo_pose_action_state_june3_1130amThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "xarm",
"total_episodes": 180,
"total_frames": 47370,
"total_tasks": 1,
"total_videos": 540,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 20,
"splits": {
"train": "0:180"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Anas0711/orange_bobo_pose_action_state_june3_1130am.UR7e-CaP-Stack_Block-100epi_10fps_state_tplus1_actionThis dataset was created using LeRobot.
Dataset Description
10fps downsampled LeRobot v3 dataset for the UR7e stack-block task. It is derived from CoRL2026-CSI/UR7e-CaP-Stack_Block-100epi by keeping source frames 0, 3, 6, ... in each clean forward-success episode. Numeric observations and absolute 7D UR7e actions are copied from retained source frames without interpolation.
Homepage: https://huggingface.co/datasets/CoRL2026-CSI/UR7e-CaP-Stack_Block-100epi
Paper: [More… See the full description on the dataset page: https://huggingface.co/datasets/Cache-SCA/UR7e-CaP-Stack_Block-100epi_10fps_state_tplus1_action.Gr00t_lerobot_state_actionThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "xarm",
"total_episodes": 181,
"total_frames": 25447,
"total_tasks": 1,
"total_videos": 362,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 10,
"splits": {
"train": "0:181"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/nimitvasavat/Gr00t_lerobot_state_action.
