datasets
Training and evaluation data, with the modality, task and licence stated up front. Listed live from the Hugging Face Hub.
vbvrpro_sampler_trajectories-baseline-steps
VBVR-Pro sampler trajectory media
This media archive backs the interactive
pufanyi/vbvrpro_sampler_trajectories
Space.
It contains 12 matched evaluation cells:
DiffSynth step-35500 baseline and DanceGRPO checkpoint 2200
Flow-CPS noise 0.1, 0.3, 0.7, and 0.9
deterministic FlowMatch Euler ODE and UniPC ODE
500 samples per cell across 100 VBVR-Pro tasks
The deployment is split across three public media repositories so each Git-backed
Dataset remains below Hugging Face's… See the full description on the dataset page: https://huggingface.co/datasets/pufanyi/vbvrpro_sampler_trajectories-baseline-steps.vbvrpro_sampler_trajectories-2200-steps
VBVR-Pro sampler trajectory media
This media archive backs the interactive
pufanyi/vbvrpro_sampler_trajectories
Space.
It contains 12 matched evaluation cells:
DiffSynth step-35500 baseline and DanceGRPO checkpoint 2200
Flow-CPS noise 0.1, 0.3, 0.7, and 0.9
deterministic FlowMatch Euler ODE and UniPC ODE
500 samples per cell across 100 VBVR-Pro tasks
The deployment is split across three public media repositories so each Git-backed
Dataset remains below Hugging Face's… See the full description on the dataset page: https://huggingface.co/datasets/pufanyi/vbvrpro_sampler_trajectories-2200-steps.VBVR-Bench-Data
VBVR: A Very Big Video Reasoning Suite
Overview
Video reasoning grounds intelligence in spatiotemporally consistent visual environments that go beyond what text can naturally capture,
enabling intuitive reasoning over motion, interaction, and causality. Rapid progress in video models has focused primarily on visual quality.
Systematically studying video reasoning and its scaling behavior suffers from a lack of… See the full description on the dataset page: https://huggingface.co/datasets/Video-Reason/VBVR-Bench-Data.vbvrpro_sampler_trajectories-data
VBVR-Pro sampler trajectory media
This media archive backs the interactive
pufanyi/vbvrpro_sampler_trajectories
Space.
It contains 12 matched evaluation cells:
DiffSynth step-35500 baseline and DanceGRPO checkpoint 2200
Flow-CPS noise 0.1, 0.3, 0.7, and 0.9
deterministic FlowMatch Euler ODE and UniPC ODE
500 samples per cell across 100 VBVR-Pro tasks
The deployment is split across three public media repositories so each Git-backed
Dataset remains below Hugging Face's… See the full description on the dataset page: https://huggingface.co/datasets/pufanyi/vbvrpro_sampler_trajectories-data.VBVR-Bench-Data
VBVR: A Very Big Video Reasoning Suite
Overview
Video reasoning grounds intelligence in spatiotemporally consistent visual environments that go beyond what text can naturally capture,
enabling intuitive reasoning over motion, interaction, and causality. Rapid progress in video models has focused primarily on visual quality.
Systematically studying video reasoning and its scaling behavior suffers from a lack of… See the full description on the dataset page: https://huggingface.co/datasets/abs794/VBVR-Bench-Data.VBVR-Bench
VBVR-Bench
Re-hosted copy of Video-Reason/VBVR-Bench-Data,
converted to standard HuggingFace parquet format.
Splits
in_domain: 50 tasks x 5 samples = 250 entries (tasks overlap with the VBVR training set).
out_of_domain: 50 tasks x 5 samples = 250 entries (held-out reasoning tasks).
Schema
field
type
notes
task_name
string
e.g. G-13_grid_number_sequence_data-generator
video_idx
string
zero-padded sample id (00000..00004)
domain
string… See the full description on the dataset page: https://huggingface.co/datasets/pufanyi/VBVR-Bench.vbvr-rl-eval-resultsvbvr-natural-2500
VBVR Natural — Current selected videos
This repository contains the current selected versions and owner-feedback revisions replacing the earlier
2,500-video campaign. The old campaign videos and their metadata have been removed
from the current dataset. The repository remains private.
Available: 500 / 500 selected videos across 100 families.
The feedback revision scope is 77 families with five samples each:
385 / 385 revised clips are available.
An additional 115 / 115 clips… See the full description on the dataset page: https://huggingface.co/datasets/xvreason/vbvr-natural-2500.VBVR-Test-Bench
VBVR-Test-Bench
A reorganization of the Video-Reason/VBVR-Bench-Data test set, organized first by split (In-Domain / Out-of-Domain), then by class (Instruction_Following / Pure_Reasoning), and inside Pure_Reasoning, further split into unchanged vs rewritten.
Class / bucket
Tasks
Samples
Instruction_Following
56
280
Pure_Reasoning / unchanged
27
135
Pure_Reasoning / rewritten
17
85
Total
100
500
rewritten tasks are reasoning tasks whose upstream VBVR prompt… See the full description on the dataset page: https://huggingface.co/datasets/May-apple/VBVR-Test-Bench.vbvrpro_output-data
VBVR-Pro strict-sweep scored videos
This media archive backs the interactive
pufanyi/vbvrpro_output
Space.
It contains 10,500 scorer-input MP4 files:
20 DanceGRPO results: 5 checkpoints × 4 sampling modes × 500 samples
1 SFT epoch-1 baseline: 500 samples
100 EvalKit-supported tasks, with 5 samples per task and run
1024×1024, 161-frame H.264 videos encoded at 33 FPS
Paths follow:
videos/{run_id}/{domain_folder}/{task_name}/{video_idx}.mp4
The Space contains the task catalog… See the full description on the dataset page: https://huggingface.co/datasets/pufanyi/vbvrpro_output-data.
