DanBenAmi/HERBench
HERBench: A Benchmark for Multi-Evidence Integration in Video Question Answering A challenging benchmark for evaluating multi-evidence integration capabilities of vision-language models ๐ HERBench has been accepted to CVPR 2026! ๐ New: Lite-v2 config. We released a refreshed lite_v2 version of the Lite split (1,971 questions / 68 videos) in which 9 of the 12 tasks were regenerated and went through additional manual refinement for higher quality, while TSO, SVAโฆ See the full description on the dataset page: https://huggingface.co/datasets/DanBenAmi/HERBench.
This repository belongs to DanBenAmi on Hugging Face.
CoolFace never edits a repository it does not host. Visibility, licence, collaborators and gating are all managed at the source.
