datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
motionatlas-data
MotionAtlas-Data
MotionAtlas-Data is a large-scale dataset for region-aware motion captioning. Instead of describing a whole clip globally, each sample pairs a video with a spatiotemporal region and a precise description of the motion inside that region, reducing visual clutter and motion entanglement.
159K high-quality region-level motion captioning samples
Built with a scalable pipeline using self-bootstrap refinement to suppress fine-grained hallucinations
Designed to… See the full description on the dataset page: https://huggingface.co/datasets/maxLWSv2/motionatlas-data.ImageNet-C-motion_blur-severity_5MotionEdit-Trainminecraft-motion-action-datasetLIBERO-motionminecraft-motion-coa-datasetomx_f_motion_1This dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "omx_f",
"total_episodes": 30,
"total_frames": 18523,
"total_tasks": 1,
"total_videos": 30,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:30"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/kdm111/omx_f_motion_1.eef_motionThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 20,
"features": {
"observation.images.robot0_eye_in_hand": {
"dtype": "video",
"shape": [
256,
256,
3
],
"names": [
"height",
"width",
"channel"
],
"video_info": {… See the full description on the dataset page: https://huggingface.co/datasets/euijinrnd/eef_motion.MotionEdit-BenchMotionEdit-Bench is a benchmark dataset for evaluating image editing models on the task of Motion Image Editing, a novel text-based image editing task that aims at modifying actions, poses, and interactions of subjects and objects in images instead of just static features like color.
You can use huggingface datasets to read our dataset:
from datasets import load_dataset
dataset = load_dataset("elaine1wan/MotionEdit-Bench")["train"]
primitive_motionThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so101",
"total_episodes": 150,
"total_frames": 3338,
"total_tasks": 30,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 10,
"splits": {
"train": "0:150"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/eorikun/primitive_motion.motion_x_rendered-bootstapir_checkpoint_v2motion_x_renderedImageNet-C-motion_blur-severity_4paramount_motionmotion_x_video-bootstapir_checkpoint_v2-crop_still_edgeseef_motion_gripper_pauseThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"fps": 20,
"features": {
"observation.images.robot0_eye_in_hand": {
"dtype": "video",
"shape": [
256,
256,
3
],
"names": [
"height",
"width",
"channel"
],
"video_info": {… See the full description on the dataset page: https://huggingface.co/datasets/euijinrnd/eef_motion_gripper_pause.cleverer-bootstapir_checkpoint_v2-crop_still_edges-discard_camera-motionclr_motion_planning_hwGitruck-MotionIR
Gitruck MotionIR
Gitruck MotionIR is a Chinese motion-design dataset that aligns project-level
natural-language descriptions, technique-level annotations, temporal evidence,
and a renderable intermediate representation (IR v1). The corpus was normalized
from authorized Alight Motion, After Effects, NodeVideo, and Jianying projects.
Gitruck MotionIR 是一个中文动效设计数据集,将工程级描述、技法级标注、时间证据与可渲染
IR v1 对齐。语料由已获授权的 Alight Motion、After Effects、NodeVideo 与剪映工程归一化而来。
Dataset summary /… See the full description on the dataset page: https://huggingface.co/datasets/Hocassian/Gitruck-MotionIR.primitive_motion_phaseThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so101",
"total_episodes": 150,
"total_frames": 3338,
"total_tasks": 30,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 10,
"splits": {
"train": "0:150"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/eorikun/primitive_motion_phase.snpmg_cot_qwen35ba3breasoning_qwen_35ba3text-2-video-human-preferences-motion
Human Preferences for AI-Generated Video: Motion Quality
29,283 pairwise human preference labels comparing 4 frontier video generation models on human motion across 3 quality dimensions, collected from 4,349 real annotators via Datapoint AI.
This is the largest publicly available human preference dataset focused specifically on human motion in AI-generated video.
Why This Dataset
Video generation models are improving fast, but evaluating human motion remains… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-video-human-preferences-motion.text-2-video-human-preferences-motion
Human Preferences for AI-Generated Video: Motion Quality
29,283 pairwise human preference labels comparing 4 frontier video generation models on human motion across 3 quality dimensions, collected from 4,349 real annotators via Datapoint AI.
This is the largest publicly available human preference dataset focused specifically on human motion in AI-generated video.
Why This Dataset
Video generation models are improving fast, but evaluating human motion remains… See the full description on the dataset page: https://huggingface.co/datasets/nusdufv/text-2-video-human-preferences-motion.d1_motionThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v2.1",
"robot_type": "Unitree_D1",
"total_episodes": 4,
"total_frames": 2377,
"total_tasks": 1,
"total_videos": 4,
"total_chunks": 1,
"chunks_size": 1000,
"fps": 30,
"splits": {
"train": "0:4"
},
"data_path": "data/chunk-{episode_chunk:03d}/episode_{episode_index:06d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/Annena/d1_motion.motionlabs-fineweb-ultra-minimotion_prediction
Visualization of Motion Prediction Task Cases Samples
Check dataset samples visualization by viewing Dataset Viewer.
The sampling procedure is guided by the Elo distribution introduced in our method.
Original dataset is validation split of Waymo Open Motion Dataset (WOMD).
samples/origin: 4409/ 44097
License
This repository is licensed under the Apache License 2.0
eval_so101_first_motionThis dataset was created using LeRobot.
Dataset Structure
meta/info.json:
{
"codebase_version": "v3.0",
"robot_type": "so_follower",
"total_episodes": 1,
"total_frames": 1420,
"total_tasks": 1,
"chunks_size": 1000,
"data_files_size_in_mb": 100,
"video_files_size_in_mb": 200,
"fps": 30,
"splits": {
"train": "0:1"
},
"data_path": "data/chunk-{chunk_index:03d}/file-{file_index:03d}.parquet",
"video_path":… See the full description on the dataset page: https://huggingface.co/datasets/0xGUI/eval_so101_first_motion.UCF-101-bootstapir_checkpoint_v2-still_edge_crop-discard_camera_motiontext-2-video-human-preferences-motion-v2-large
Human Preferences for AI-Generated Video: Motion Quality v2 (large)
115,732 pairwise human preference labels comparing 4 frontier video generation models on human motion across 3 quality dimensions, collected from real annotators via Datapoint AI.
This is an expanded version of the motion quality dataset with 417 unique prompts (up from 60) and 11 motion categories (up from 6).
Why This Dataset
Video generation models are improving fast, but evaluating human motion… See the full description on the dataset page: https://huggingface.co/datasets/datapointai/text-2-video-human-preferences-motion-v2-large.
