CoolFace
Datasetpublic

DanBenAmi/HERBench

HERBench: A Benchmark for Multi-Evidence Integration in Video Question Answering A challenging benchmark for evaluating multi-evidence integration capabilities of vision-language models 🎉 HERBench has been accepted to CVPR 2026! 🆕 New: Lite-v2 config. We released a refreshed lite_v2 version of the Lite split (1,971 questions / 68 videos) in which 9 of the 12 tasks were regenerated and went through additional manual refinement for higher quality, while TSO, SVA… See the full description on the dataset page: https://huggingface.co/datasets/DanBenAmi/HERBench.

sourceHugging Facecc-by-nc-sa-4.0updated 4mo agoView on Hugging Face
3likes1.8kdownloads
discussions and pull requests

Conversations for this repository live on Hugging Face.

CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.

Open discussions on Hugging Face