macrodata/WGO-Bench
WGO-Bench: What's Going On Benchmark WGO-Bench is a small, manually annotated benchmark for evaluating how well vision-language models can turn robot and egocentric manipulation videos into timestamped subtask annotations. Each row contains one video episode, a high-level task instruction, and gold subtask segments with start time, end time, and a concise action label. The benchmark is designed for two related tasks: Boundary detection: predict where one meaningful manipulation… See the full description on the dataset page: https://huggingface.co/datasets/macrodata/WGO-Bench.
This repository belongs to macrodata on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
