Alibaba-NLP/xvbench
XVBench XVBench is a benchmark for evaluating multimodal retrieval-augmented generation systems on cross-video understanding. It is introduced alongside the paper VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph. The questions in XVBench are created based on videos from HowTo100M, a large-scale corpus of narrated instructional videos. The benchmark focuses on questions that require models or agents to retrieve and reason… See the full description on the dataset page: https://huggingface.co/datasets/Alibaba-NLP/xvbench.
Conversations for this repository live on Hugging Face.
CoolFace shows imported repositories read-only. Posting into someone else’s repository from here would need an authorised integration and the account holder’s consent, so the link goes to the source instead.
Open discussions on Hugging Face